Review
Open access
Jul 2026
Unboxing the Black Box: A Survey on Mechanistic Interpretability for Algorithmic Understanding of Neural Networks
It is argued that mechanistic interpretability has the potential to support a more scientific understanding of machine learning systems – treating models not only as tools for solving tasks, but also as systems to be studied and understood.
Bianka Kowalska, Halina Kwasnicka
· Machine-mediated learning · 0 citations