Skip to content
Open access

MSstatsBioNet: Integrating Statistical Analyses with Prior Knowledge Biomolecular Networks for Quantitative Proteomics and Phosphoproteomics

Jul 2026 · bioRxiv · 0 citations
Biology

TL;DR

This manuscript automates the integration of biological network databases with MSstatsBioNet, a Bioconductor package that integrates MSstats, a family of open-source packages for detecting differentially abundant proteins, and INDRA, a system that extracts biomolecular networks from biomedical literature using text mining and merges those networks with the content of curated knowledge bases.

Abstract

A common outcome of quantitative mass spectrometry-based proteomic and phosphoproteomic experiments is a list of proteins that are differentially abundant between conditions. However, biological interpretation requires evaluation in the context of prior knowledge of biological mechanisms and protein function. One approach to facilitate mechanistic biological interpretation is to integrate such lists with biological network databases, built from manually curated resources and text mining systems. This manuscript automates this process with MSstatsBioNet, a Bioconductor package that integrates MSstats, a family of open-source packages for detecting differentially abundant proteins, and INDRA, a system that extracts biomolecular networks from biomedical literature using text mining and merges those networks with the content of curated knowledge bases. Taking as input a list of differentially abundant proteins from MSstats, MSstatsBioNet retrieves a protein subnetwork from INDRA and overlays experimental fold changes onto the underlying subnetwork. Users can then interact with the network and overlaid data, interrogating primary literature evidence to construct granular mechanistic narratives for iterative hypothesis generation. We demonstrate the utility of this approach with three case studies, two measuring changes in protein abundance and one measuring changes in phosphorylation.

Read PDF

Similar papers

Open access Oct 2026

PiProteline: An R Package for Integrated Proteomics Data Analysis, from Label-Free Quantitation to Systems Biology

The demand for user-friendly applications to support biologists in analyzing high-throughput proteomics data remains a pressing challenge. Given the complexity and the multiple intermediate steps involved, this process is time-consuming and often requires specialized computational skills. To simplify and accelerate the...

Andrea Lomagno, S. Hamed, Ishak Yusuf et al. · 0 citations
#large language models Open access Sep 2026

Atlantis: An integrative database for human proteome structural and functional sites

Atlantis, a database that integrates structural and functional information at the human proteome residue level and a Model Context Protocol (MCP) connector allows the interrogation of the resource through Large Language Models (LLMs) or agentic frameworks for biomedical research.

Natalia De Oliveira Rosa, Piergiorgio Ferronato, M. Varisco et al. · 0 citations
2026

Computational and Statistical Framework for Quantitative Proteomics Analysis.

This chapter presents a step-by-step pipeline for the statistical and computational analysis of such data, oriented and generalizable to any mass spectrometry-derived proteomic dataset, facilitating an end-to-end analysis from raw proteomic data to the biological interpretation.

Ismail Kirrout, N. Montes · 0 citations
Open access Sep 2026

ProteoformTracker: an interactive tool for planning proteoform detectability in top-down and middle-down proteomics

Summary Characterizing proteome complexity in disease contexts is essential for understanding molecular mechanisms and advancing therapeutic development. Mass spectrometry (MS)-based top-down and middle-down proteomics (TDP/MDP) can resolve intact proteoforms — protein molecules carrying a unique combination of isoform...

Araf Mahmud, Zhi-Hao Zhang, Si Wu et al. · 0 citations
Open access Aug 2026

ProtPen Combines Sequence- and Structure-based Approaches to Facilitate Protein Function Predictions on a Proteome-wide Scale.

Proteins of unknown function represent a significant gap in our understanding of biological processes, encompassing large portions of the proteomes of many organisms, especially prokaryotes. Addressing this gap is critical to understanding the biology and pathogenicity of such organisms. We introduce ProtPen, an open-s...

Diya Mathai, S. Schulze · 0 citations
Open access Aug 2026

Sample-specific protein-protein interaction networks inferred from transcriptomics and proteomics show high similarities

Contextualized protein-protein interaction networks provide crucial insight into diseases and other biological processes, but for a profound understanding of such processes and their distinct effects on individuals, the protein-protein interactions within individual samples must be investigated. A straightforward appro...

Enikő Zakar-Polyák, C. Kerepesi · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.