MEGA Hub

Traceable Spectral Inference via Influence Functions: Efficient Data Attribution and Error Proxies for the Ariel Mission

Authors

Do you know Nikki Grens?You can claim authorship or link another user.Do you know Luís F. Simões?You can claim authorship or link another user.Do you know Kai Hou Yip?You can claim authorship or link another user.Do you know Theresa Lueftinger?You can claim authorship or link another user.

Abstract

Interpretability is critical for machine learning models deployed in scientific space missions such as ESA's Ariel, where ground truth is unavailable during operations and physical plausibility must be assessed. While most explainable AI methods focus on feature attribution, this work investigates training data attribution through influence functions and introduces three key contributions for operational spectroscopy pipelines. First, influence is reformulated in terms of prediction rather than loss, enabling label-free deployment. Second, by leveraging the closed-form ridge solution of an Extreme Learning Machine, infinitesimal prediction influence is efficiently computed. Third, an influence-based conservative error proxy is derived by propagating training residuals through the influence sensitivities. Evaluated against simulated spectra, the proposed proxy correlates strongly with scale and shape-based spectral errors. Furthermore, influence functions enable the identification of the most influential samples and the approximation of the most harmful ones. Together, these results suggest that this approach can serve as an operational framework for scientific machine learning.

Community

00

Publication notes

Author note
To appear in "Proceedings of SPAICE 2026: Third Conference on AI in and for Space"