MultiSpaceNet is presented, a graph-based framework for joint representation learning from paired spatial transcriptomic and proteomic data that outperforms nine published state-of-the-art methods in spatial domain identification and preserves biological structure across replicate sections better than all compared alternatives.
Abstract
Abstract Motivation Paired spatial assays now measure the transcriptome and the proteome on the same tissue section. Existing integration methods commit to a structural choice at the outset: a separate graph per modality, early fusion that folds the measured protein into the embedding so it can no longer be predicted, or intersection rules that discard most similarity edges. Each choice limits what one trained model can do afterwards. Results We present MultiSpaceNet, a graph-based framework for joint representation learning from paired spatial transcriptomic and proteomic data. A section is represented as one cell graph over a shared node set, with spatial, transcriptomic and proteomic relations carried as typed edges and the two modality branches kept separate until a per-cell attention fusion. On five benchmark datasets, MultiSpaceNet outperforms nine published state-of-the-art methods in spatial domain identification (mean adjusted Rand index 0.472). Its leakage-free RNA-to-protein imputation matches or exceeds established methods for signal-bearing proteins, and it preserves biological structure across replicate sections better than all compared alternatives. A single trained model thus provides spatial domains, protein imputation, cross-section joint embedding and descriptive per-cell modality-dominance maps. Availability and implementation Source code is available at https://github.com/yongzhuangliulab/MultiSpaceNet and archived at Zenodo (DOI 10.5281/zenodo.22667390 and 10.5281/zenodo.21488824). The scripts that regenerate the reported tables and figures, together with their machine-readable result summaries, are included in the repository (directory resubmission/) and its Zenodo archive.
Mapping spatially coherent tissue domains from spatial transcriptomics data is a prerequisite for characterizing cell-type, developmental patterning, and disease-associated disruptions of tissue architecture. While graph neural network (GNN) methods have substantially improved domain identification over expressio...
M. Al-Taie, Firas Hazzaa, Akram Qashou et al.· BMC Bioinformatics· 0 citations
Abstract Paired spatial multi-omics provides a supervised basis for learning RNA–protein correspondence in situ, but predicting protein abundance from spatial transcriptomic data alone remains challenging across tissue contexts and protein panels. Here, we present DPAS-Graph, an adaptive relation-learning framework for...
The rapid development of Spatial Transcriptomics (ST) enables simultaneous acquisition of gene expression and spatial locations, offering new avenues to explore tissue organization. However, effectively integrating spatial and transcriptional information for spatial domain identification remains challenging in existing...
Wei Zhang, Dan-Yang Dong, Bang-Yi Zhang et al.· IEEE transactions on computa...· 0 citations
Single-cell RNA sequencing (scRNA-seq) profiles transcriptomes at high resolution but discards the spatial context of cells within a tissue—information that is essential for studying intercellular mechanisms and tissue architecture. Spatial transcriptomics (ST) retains coordinates but, depending on the assay, trades th...
Sebastian Birk, Fabian J. Theis, M. Lotfollahi· bioRxiv· 0 citations
Foundation models offer a promising paradigm for modeling spatial transcriptomics, but capturing tissue context over cellular graphs makes training at scale challenging. We introduce spaGFM, a graph foundation model that serializes cellular neighborhoods through random walks to generate transformer-compatible represent...
Yu Zhong, Fei He, Xiao-Jie Jin et al.· Research Square· 0 citations
Results support graph-connected gene blocks as useful prediction units for JEPA-style representation learning in single-cell biology by supporting block-level prediction of graph-connected gene blocks defined by protein-association and corpus-derived coexpression evidence.
Yu-Hao Wang, Ze-Lin Zang, Yuxuan Liu et al.· 0 citations
A new machine-learning framework aims to improve the success rate of computational protein design while moving away from results that reproduce sequences found in nature.