Skip to content
Review Open access

Artificial intelligence for genomic science: a scoping review of concepts, architectures, applications, and open challenges

Jul 2026 · Frontiers in Bioinformatics · Vol 6 · 1 citation · 54 references
Medicine

TL;DR

This scoping review mapped how AI is defined and operationalized in genomic science, including machine learning, deep learning, graph-based methods, foundation models, and large language models, and synthesized their data modalities, applications, evaluation practices, interpretability strategies, and governance challenges.

Abstract

Introduction Artificial intelligence (AI) is becoming central to genomics and multi-omics, but its concepts, architectures, applications, evaluation standards, and translational requirements remain fragmented. This scoping review mapped how AI is defined and operationalized in genomic science, including machine learning, deep learning, graph-based methods, foundation models, and large language models, and synthesized their data modalities, applications, evaluation practices, interpretability strategies, and governance challenges. Methods We conducted a PRISMA-ScR scoping review with Joanna Briggs Institute guidance. Eligible studies applied AI to genomics or closely allied omics in research, clinical, or public health contexts. MEDLINE/PubMed, Embase, and supplementary registers were searched from January 2001 to 3 September 2025 without language restrictions. Records were screened in duplicate, and standardized items were extracted, including AI concept or method family, omics modality, task, metrics, interpretability, governance, and deployment considerations. Methodological reporting and quality were appraised using design-appropriate JBI tools and summarized descriptively as a normalized 0%–100% checklist-fulfillment index. Results From 3,785 records, 1,040 studies were included. Publication remained sparse until 2017 and then expanded steeply, with more than 90% appearing from 2018 onward. The normalized JBI checklist-fulfillment index was modest overall (mean 35.3%, SD 20.1; range 7.5%–87.5%) and was interpreted descriptively, not as a directly comparable quality score across designs. Conceptually, the field has moved from feature-engineered statistical learning toward representation learning systems modeling nucleotide sequences, regulatory context, single-cell states, multi-omics profiles, biomedical text, and clinical-genomic knowledge. Applications concentrated on variant interpretation, regulatory genomics, multi-omics integration, single-cell analysis, pathology/radiology-genomics fusion, and genomic decision support, with increasing use of deep learning, graph models, foundation models, and LLMs. Calibration, external validation, mechanistic interpretability, ancestry-aware fairness, privacy protection, and deployment models for sensitive genomic data were unevenly reported; prospective multisite evaluations were rare. Discussion AI in genomics has scaled rapidly since 2017–2018, but translation remains constrained by heterogeneous concepts, inconsistent benchmarks, incomplete reporting, and limited governance. Priorities include biologically meaningful benchmarks; calibrated uncertainty for genomic decision support; mechanism-linked interpretability; ancestry- and site-aware validation; privacy-preserving analysis of sensitive genomic data; and human oversight for variant interpretation, precision medicine, and public health genomics. Systematic Review Registration https://osf.io/uexzh.

Read PDF

Similar papers

#artificial intelligence Review Aug 2026

Responsible Integration of AI in Cancer Genomics: Barriers, Risks, and Pathways to Trustworthy Clinical Translation

Artificial intelligence (AI) and natural language processing (NLP) are increasingly used to extract, integrate, and interpret biomedical knowledge relevant to cancer genomics, yet their translation into routine clinical oncology has been comparatively slow. The central challenge is not computational capability alone, but trustworthy integration into clinical workflows. This review examines how NLP and AI support the cancer genomics pipeline, from literature mining and automated variant interpretation to clinical trial matching, knowledge graph construction, and multimodal data integration. We identify four interrelated translational failure domains: evidence inconsistency, explainability and uncertainty, data governance and reproducibility, and interoperability. Rather than considering these challenges in isolation, we take a systems-level view, focusing on their interaction across the translational pathway. We propose a conceptual framework and roadmap for addressing these domains through rigorous validation, uncertainty-aware methods, interoperable infrastructures, regulatory alignment, and human oversight across the AI lifecycle. Progress toward routine clinical use will depend less on further improving model capability than on systematically addressing these interacting failure domains from development through deployment and post-deployment monitoring.

B. Ilgen, Yiannos S. Tolias, Denise Kühnert et al. · 0 citations
Review Open access Jul 2026

Artificial Intelligence and Genomic Data Analysis: New Frontiers in Precision Medicine

A clinically oriented, pipeline-based synthesis of contemporary AI applications in genomic medicine, focusing on factors that determine model robustness and clinical utility, and common sources of failure in real-world genomic AI systems.

Alexandra-Maria Blaga, Răzvan-Octavian Mihuț, A. Treteanu et al. · 0 citations
Review Open access Aug 2026

Integrating Artificial Intelligence with Global Genomic Resources: A Narrative Review of Implications for Precision Medicine

Findings indicate that artificial intelligence can support the integration and analysis of multi-omics data, support the identification of genetic variants and disease associations, and improve predictive modeling for precision medicine.

Towsif Alam, Koushik Saha, M. K. K. Rony et al. · 0 citations
Review Open access Jul 2026

Generative AI in Healthcare: Applications and Challenges

Generative AI demonstrably accelerates diagnostic workflows, augments scarce clinical datasets, personalizes communication, and supports discovery pipelines, and the paper concludes with a translational path and research priorities aimed at closing these gaps.

Wael Rahhal · 0 citations
Review Open access Aug 2026

ARTIFICIAL INTELLIGENCE IN MODERN MEDICINE: A COMPREHENSIVE REVIEW OF CURRENT APPLICATIONS, CHALLENGES, AND FUTURE PERSPECTIVES

Introduction: The digital transformation of healthcare is accelerating, driven by unprecedented advancements in Artificial Intelligence (AI). From large language models (LLMs) to biomolecular structure prediction, AI is redefining modern diagnostic and therapeutic standards. Aim: This review evaluates the current state of AI applications in medicine, focusing on clinical knowledge encoding, molecular drug discovery, and administrative workflow optimization, while critically addressing the technical, ethical, and systemic challenges of their institutional implementation. Materials and Methods: A structured analysis was conducted utilizing a hybrid approach that combines a multi-decade bibliometric trend perspective with a detailed synthesis of 21 landmark publications, clinical trials, and meta-analyses from high-impact journals. Results: AI demonstrates expert-level performance in medical knowledge retrieval and spatiotemporal diagnostics. AlphaFold 3 has revolutionized computational therapeutics through all-atom biomolecular interaction prediction, while ambient AI scribes significantly reduce physician burnout by automating clinical documentation workflows. However, data-driven "hallucinations" in LLMs and the inherent "black box" nature of deep learning architectures remain critical barriers to autonomous deployment. Conclusions: AI is successfully transitioning from an isolated research tool into an essential clinical "co-pilot." Achieving its full potential in Medicine 4.0 requires robust frameworks for algorithmic explainability, global dataset diversification, and a strategic synergy between machine precision and human clinical judgment.

Aleksandra Stańczyk, Kinga Haduch, Zuzanna Michalska et al. · 0 citations
Review Jul 2026

Agentic AI in medicine: architectures, applications, evaluation, and challenges for clinical translation

Large language models and multimodal foundation models are enabling medical artificial intelligence (AI) systems to move beyond isolated prediction and undertake multistep clinical tasks that require planning, tool use, memory, iterative correction, and coordination among specialized agents. However, the scope of agentic AI in medicine remains unsettled, and current evaluation practices are not yet aligned with the requirements of clinical use. We conducted a scoping review with systematic evidence mapping across five electronic sources, screened 1,649 exportable records, and provisionally included 557 unique studies that met predefined criteria for goal-directed task execution, tool use, interaction with external resources, feedback-based refinement, or multi-agent collaboration. The included studies describe single agents that use external tools, workflows supported by retrieval and external knowledge, multimodal agents, and multi-agent systems applied to medical question answering, image interpretation, electronic health record analysis, drug safety, and clinical trial prediction. The evidence base remains dominated by public benchmarks, simulated settings, retrospective datasets, and small-scale expert evaluation. Process reliability, evidence traceability, uncertainty, safety, workflow impact, and external validity are evaluated less consistently. Clinical translation will depend on clearer definitions, reproducible evaluation, auditable oversight, interoperable system design, and prospective validation in real-world clinical workflows.

Zheng Tong, Yang Liu, Wanshu Fan et al. · 0 citations