How Interpretable Machine Learning Can Benefit Process Understanding in the Geosciences

Shijie Jiang; Lily belle Sweet; Georgios Blougouras; Alexander Brenning; Wantong Li; Markus Reichstein; Joachim Denzler; Wei Shangguan; Guo Yu; Feini Huang; Jakob Zscheischler

doi:10.1029/2024EF004540

How Interpretable Machine Learning Can Benefit Process Understanding in the Geosciences

Publikation: Beitrag in Fachzeitschrift › Forschungsartikel › Beigetragen › Begutachtung

Beitragende

Shijie Jiang - , Max Planck Institute for Biogeochemistry, European Laboratory for Learning and Intelligent Systems (Autor:in)
Lily belle Sweet - , Professur Data Analytics in Hydro Sciences (gB/UFZ), Helmholtz-Zentrum für Umweltforschung (UFZ) (Autor:in)
Georgios Blougouras - , Max Planck Institute for Biogeochemistry, European Laboratory for Learning and Intelligent Systems, Friedrich-Schiller-Universität Jena (Autor:in)
Alexander Brenning - , European Laboratory for Learning and Intelligent Systems, Friedrich-Schiller-Universität Jena (Autor:in)
Wantong Li - , Max Planck Institute for Biogeochemistry (Autor:in)
Markus Reichstein - , Max Planck Institute for Biogeochemistry, European Laboratory for Learning and Intelligent Systems (Autor:in)
Joachim Denzler - , European Laboratory for Learning and Intelligent Systems, Friedrich-Schiller-Universität Jena (Autor:in)
Wei Shangguan - , Sun Yat-Sen University (Autor:in)
Guo Yu - , Desert Research Institute (Autor:in)
Feini Huang - , Max Planck Institute for Biogeochemistry, European Laboratory for Learning and Intelligent Systems, Sun Yat-Sen University (Autor:in)
Jakob Zscheischler - , Center for Scalable Data Analytics and Artificial Intelligence (ScaDS.AI Dresden), Professur Data Analytics in Hydro Sciences (gB/UFZ), Helmholtz-Zentrum für Umweltforschung (UFZ) (Autor:in)

Abstract

Interpretable Machine Learning (IML) has rapidly advanced in recent years, offering new opportunities to improve our understanding of the complex Earth system. IML goes beyond conventional machine learning by not only making predictions but also seeking to elucidate the reasoning behind those predictions. The combination of predictive power and enhanced transparency makes IML a promising approach for uncovering relationships in data that may be overlooked by traditional analysis. Despite its potential, the broader implications for the field have yet to be fully appreciated. Meanwhile, the rapid proliferation of IML, still in its early stages, has been accompanied by instances of careless application. In response to these challenges, this paper focuses on how IML can effectively and appropriately aid geoscientists in advancing process understanding—areas that are often underexplored in more technical discussions of IML. Specifically, we identify pragmatic application scenarios for IML in typical geoscientific studies, such as quantifying relationships in specific contexts, generating hypotheses about potential mechanisms, and evaluating process-based models. Moreover, we present a general and practical workflow for using IML to address specific research questions. In particular, we identify several critical and common pitfalls in the use of IML that can lead to misleading conclusions, and propose corresponding good practices. Our goal is to facilitate a broader, yet more careful and thoughtful integration of IML into Earth science research, positioning it as a valuable data science tool capable of enhancing our current understanding of the Earth system.

Details

Originalsprache	Englisch
Aufsatznummer	e2024EF004540
Fachzeitschrift	Earth's Future
Jahrgang	12
Ausgabenummer	7
Publikationsstatus	Veröffentlicht - Juli 2024
Peer-Review-Status	Ja

Schlagworte

ASJC Scopus Sachgebiete

Schlagwörter

big data, interpretability, interpretable machine learning, knowledge discovery, machine learning, XAI