Geology ReportsSearch

USGS · 70273156

Achieving interpretable machine learning by functional decomposition of black-box models into explainable predictor effects

Abstract

Machine learning (ML) models are often based on complex black-box architectures that are difficult to interpret. This interpretability problem can hinder the use of ML in fields like medicine, ecology, and insurance, and has boosted research in interpretable machine learning (IML). Here, we propose a novel approach for the functional decomposition of black-box predictions, which is a core concept of IML. This approach replaces the prediction function with a surrogate model consisting of simpler subfunctions, providing insights into the direction and strength of the main feature contributions and their interactions. Our method is based on a concept termed “stacked orthogonality”, which ensures that the main effects capture as much functional behavior as possible. To compute the subfunctions, we combine neural additive modeling with an efficient post-hoc orthogonalization procedure. Our method yielded plausible results in an analysis of stream biological condition in the Chesapeake Bay watershed (United States).

Explore related subjects

90° N90° S · 180° W ← longitude → 180° E
Source-reported bounding extent: 36.48314061639213° to 42.98053954751642° latitude; -80.771484375° to -74.278564453125° longitude. This indicates report coverage, not an exact sampling location. View area on OpenStreetMap.

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

David Kohler, David Rügamer, Lindsey J. Boyle, Kelly O. Maloney, Matthias Schmid. 2025-11-03. Achieving interpretable machine learning by functional decomposition of black-box models into explainable predictor effects. https://doi.org/10.1038/s44387-025-00033-7

Cite the original work for its findings. Save a collection to share your selection of sources.