FACTMx couples latent patient factors with subobservation clustering and per-patient component proportions, enabling direct interpretation and downstream association analyses, and supports joint structured-simple modelling for interpretable multimodal patient stratification.
Abstract
Patient cohort profiling increasingly includes structured views for multiple modalities, such as single-cell RNA sequencing, spatial transcriptomics or proteomics, and histology, each providing multiple subobservations per patient, including single cells, spatial spots or patches. To model such data along with simple patient-level views, current multimodal integration methods typically rely on separately precomputed summaries and fail to fully leverage information in structured views. Here we present FACTMx, a variational framework that jointly models structured and simple views to learn interpretable patient-level representations. FACTMx couples latent patient factors with subobservation clustering and per-patient component proportions, enabling direct interpretation and downstream association analyses. The framework supports different structured-view mixture assumptions, including topic- and Gaussian-structured data, while retaining modular encoder-decoder parameterisations. In simulations spanning sparse and dense dependencies and multiple noise regimes, FACTMx improved reconstruction, integration and recovery of structured components relative to previous methods. Applied to non-small cell lung cancer cohorts, FACTMx captured survival-associated latent signals linked to immune microenvironments, gene expression pathways and spatially coherent histological patterns. In a longitudinal coronary syndrome cohort, FACTMx highlighted an outcome-associated axis connected to ejection-fraction change, immune cell states, soluble mediators and cardiac injury markers. These results support joint structured-simple modelling for interpretable multimodal patient stratification.
This work proposes MultiSigBERT, a unified framework for multimodal sequential survival modeling in oncology based on path signature representations that achieves a concordance index of 0.743 on an independent test set, demonstrating the benefit of jointly modeling multimodal temporal dynamics together with patient-lev...
Paul Minchella, Stéphane Chrétien, Guillaume Metzler et al.· 0 citations
Integrating heterogeneous clinical modalities, structured electronic health records (EHRs), clinical text, and medical imaging is crucial for reliable clinical prediction, yet real-world data are often sparse and imbalanced. Furthermore, prior approaches treat temporal dynamics and inter-patient relationships in isolat...
Shivani Gupta, H. Kumar, Joydeep Chandra· Proceedings of the Thirty-Fi...· 0 citations
Multi-omics integration has become central to precision medicine, yet in real-world clinical cohorts, complete multi-layer profiling is rarely achieved. Cost, assay failure, and evolving study design frequently produce block-wise modality missingness, where entire omics layers are absent for subsets of patients. This...
R. Nguyen, F. Vafaee· Artificial Intelligence Revi...· 0 citations
LatentVerse is a representation analysis resource that combines a web-based visual analytics platform for accessible, report-driven exploration with a command-line interface for scalable technical workflows that makes foundation model representations more understandable in biomedical and data science applications.
Majd Alafrange, S. Friedman, J. Kitonyo et al.· 0 citations
HounsBench is introduced, a computed tomography (CT) centric patient-state benchmark that unifies these three task families with patient-disjoint splits and per-family metrics, and HounsWorld, a 3B multimodal world model that treats volumetric scans and language as observations of the shared state through Joint Underst...
Yun-Hao Bai, Zhongwei Qiu, Guangyu Guo et al.· 0 citations
Multimodal single-cell assays profile complementary layers of cell state, but integration is complicated by modality mismatch, sparsity, and uneven cohort coverage. Here, we present Unified Variational Inference (UniVI), a scalable mixture-of-experts β-variational autoencoder that learns a shared latent space while pre...
Andrew J. Ashford, Trevor Enright, Julia Somers et al.· Genome Research· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.