Skip to content

KANEx: Translating Kolmogorov-Arnold Networks' Interpretability to Medical Explainability

Jul 2026 · arXiv.org · Vol abs/2607.24730 · 0 citations · 28 references
Computer Science

TL;DR

KANEx is introduced, the first ever framework that leverages the symbolic transparency of KANs to ground VLM reasoning, and suggests that grounding linguistic explanations and visual attributions in mathematically interpretable units is a necessary step toward trustworthy medical AI.

Abstract

Computer vision models have become highly effective for medical applications, yet their black-box nature continues to undermine clinician trust. In clinical workflows, chest X-ray classifiers are increasingly paired with Vision-Language Models (VLMs) to generate natural-language explanations. However, these systems add linguistic fluency without addressing the underlying opacity of the visual model. With the emergence of Kolmogorov-Arnold Networks (KANs), whose spline-based components provide inherently interpretable functional units, we investigate whether this architectural transparency can be leveraged to produce more trustworthy textual explanations. We introduce KANEx, the first ever framework that leverages the symbolic transparency of KANs to ground VLM reasoning. This interpretability also made it possible to design KAN-Map, a novel heatmap generation method derived directly from KAN models rather than gradient approximations. We feed these grounded contexts into downstream VLMs for enhanced explainability. Benchmarked on the MIMIC-CXR dataset, we demonstrate that KAN-based architectures with ResNet/ViT baselines demonstrate improved semantic similarity while producing significantly more faithful saliency maps. KAN architectures improve visual localization and downstream reasoning quality by 10%. Our findings suggest that grounding linguistic explanations and visual attributions in mathematically interpretable units is a necessary step toward trustworthy medical AI.

View source

Similar papers

Preprint Sep 2026

Transferring Visual Explanations: How Cross-Architecture Knowledge Distillation Affects Model Interpretability

Deploying efficient neural networks is essential in resource-constrained environments, yet compact models often sacrifice interpretability - a critical in safety-critical domains such as autonomous driving and medicine. This study investigates whether Knowledge Distillation transfers the spatial feature attribution of...

Aleks Czufarow, I. Babin · 0 citations
Conference Open access Sep 2026

MedVCoT: Bridging the Modality Gap in Medical VQA Through Latent Visual Reasoning

This work proposes MedVCoT, which incorporates latent visual reasoning into the medical visual question answering (VQA) domain, and utilizes the specialized expertise of MedSAM to train a large vision-language model so that it can autonomously generate consistent and continuous latent visual tokens within Visual Chain-...

Bo Xu, Quan-Hao Zhu, Bo-Lin Zhu et al. · 1 citation
Open access Aug 2026

An explainable biomedical foundation model via large-scale concept-enhanced vision-language pretraining.

Across a large-scale benchmark covering 78 datasets in 10 imaging modalities, ConceptCLIP demonstrates superior diagnostic performance while providing human-understandable explanations, and represents a critical milestone towards the widespread clinical adoption of AI.

Yuxiang Nie, Sunan He, Yequan Bie et al. · 0 citations
Open access Sep 2026

Explainable machine learning for leukaemia diagnosis: enhancing interpretability and trust

One of the major challenges in applying machine learning (ML) to leukaemia diagnosis is the limited interpretability of deep learning models such as artificial neural networks (ANNs) and convolutional neural networks (CNNs). While these models excel at detecting complex patterns in medical data, they often function as...

Hamza Abu Owida, Mohammad R. Hassan, Nidal M. Turab et al. · 0 citations
Open access Aug 2026

Towards Interpretable AI Second Opinions: Foundation Model Heatmaps in Radiology

This work developed an interactive application that enables readers to engage directly with model-generated heatmaps as they form their diagnoses, and conducted a user study to evaluate how this influences diagnostic behaviour and accuracy.

E. Dack, C. Dai, H. Hoppe et al. · 0 citations
Preprint Aug 2026

Scaling Inherently Interpretable Language Models

Steerling-8B remains competitive with open peer models trained on substantially 2-16x more compute, suggesting a different scaling paradigm: interpretability can be designed into training, and it improves with scale.

Guide Labs Team, Andreas Madsen, A. Ismail et al. · 3 citations · ⚡1

Related blog posts

Microsoft Research Blog Oct 7, 2026

Agent Lightning v1.0: A 3,500-Line Lightweight Agentic RL Framework for Training Agents with Real Harnesses

Training AI agents with reinforcement learning can be challenging because their tools, context, and decision-making are managed by complex frameworks. Agent Lightning connects existing agents to RL training, making it easier to improve them without rebuilding them. The post Agent Lightning v1.0: A 3,500-Line Lightweight Agentic RL Framework for Training Agents with Real Harnesses appeared first on Microsoft Research.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.