DeepVaris is introduced, an explainable deep learning framework that reframes feature selection as the interpretation of a pretrained convolutional neural network via surrogate modeling that will serve as a robust tool for biomarker discovery.
Abstract
Abstract Identifying essential biomarkers remains a core challenge in elucidating the pathogenic mechanisms and achieving precise diagnosis of complex diseases. Deep neural networks offer immense predictive power, yet their lack of interpretability severely limits downstream biological insight. Here, we introduce DeepVaris, an explainable deep learning framework that reframes feature selection as the interpretation of a pretrained convolutional neural network via surrogate modeling. In extensive simulations and real-world datasets, DeepVaris successfully identifies important features and reveals deeper insight into different diseases. Specifically, it overcomes extreme feature sparsity to identify crucial microbial biomarkers in preterm birth pregnancies. In single-cell RNA sequencing data, it reveals key transcriptional drivers governing myelin regeneration in neurodegenerative diseases missed by traditional differential expression analysis. Furthermore, in complex breast cancer cohorts, DeepVaris moves beyond generic pan-cancer signals to precise subtype-specific microenvironmental targets. In summary, we believe that DeepVaris will serve as a robust tool for biomarker discovery.
Biomarkers are central to modern diagnostics and therapeutics, yet traditional discovery approaches suffer from single-modality analyses, weak mechanistic foundations, and low reproducibility. Artificial intelligence (AI) enables the integration of complex multimodal biomedical data, but the translation of AI-derived b...
Raşit Dinç, N. Ardic· Signal Transduction and Targ...· 0 citations
Identifying biomarkers of cancers presents a persistent challenge due to insufficient interpretability in model decision. Current approaches only provide explainability that quantifies the contribution of input features or local substructures to predictions at the data level, yet they fail to uncover the inherent reaso...
Ping Zhang, Weicheng Sun, Jin-Sheng Xu et al.· Journal of King Saud Univers...· 0 citations
The proposed engGNN is highlighted as a robust, flexible, and interpretable framework for disease classification and biomarker discovery in high-dimensional omics contexts and provides interpretable feature importance scores that facilitate biologically meaningful discoveries, such as pathway enrichment analysis.
Tian-Tian Yang, Yuxuan Wang, Zhen-Wei Zhou et al.· Briefings in Bioinformatics· 0 citations
High-dimensional ARF (h-ARF), an extension of ARF optimized for integrated clinical and high-dimensional omics data, is introduced, showing that h-ARF better preserves both feature distributions, and downstream clustering and prediction utilities compared with ARFs.
C. Fouodo, J. Kapar, Anke Huels et al.· bioRxiv· 0 citations
An interpretable, multi-modal framework that integrates histopathological image analysis with multi-omics profiling, leveraging U-Net-based nuclei segmentation, vision-language models (BLIP), biomedical language models (BioGPT), and explainable AI is proposed.
AL Imran, Khandokar Md. Rahat Hossain, SM Rafiqul Islam et al.· bioRxiv· 0 citations
A structured multi-layer interpretation framework is proposed that links computational outputs across data-level processing, epigenomics-informed integrative regulatory modeling, and multi-omics-informed clinical interpretation, enabling traceable and mechanistically interpretable clinical inference.
M. Srivastava, Pratik Kumar, Ankita Chouhan et al.· Academia Molecular Biology a...· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.