Skip to content
Open access

Less Can Be Better: Decomposing Clinical Data Modalities in Large Language Model-based Healthcare Applications

Sep 2026 · JAMIA Journal of the American Medical Informatics Association · 0 citations
Medicine

TL;DR

The benefits of multimodal data integration are task-dependent and healthcare LLMs should examine clinical data modalities according to specific tasks for efficient integration, and provide practical guidance for designing efficient clinical decision support systems.

Abstract

Objective To systematically examine different clinical data modalities in large language models (LLMs) and multimodal large language models (MLLMs), and to quantify the contribution of data modalities in early inpatient risk prediction and decision support tasks. Materials and Methods We conducted a systematic analysis using MIMIC-IV, MIMIC-IV-Note, and MIMIC-CXR-JPG datasets to create a unified cohort of 22,254 hospital admissions containing structured electronic health records (EHRs), radiology reports (clinical notes), and chest X-ray images. We evaluated general-purpose and medical-adapted LLM/VLMs across uni-, bi-, and tri-modal configurations on two risk prediction tasks (in-hospital mortality and length-of-stay [LOS] prediction) and two clinical decision support (CDS) tasks (discharge diagnosis phenotyping and medication-use prediction). Results For risk prediction tasks, structured EHR data alone achieved the best or comparable performance (best mortality AUROC: 0.849; LOS AUROC: 0.868), with limited incremental benefit observed from adding radiology reports or medical images. For CDS tasks, multimodal integration yielded substantial improvements: the best tri-modal configuration achieved F1-scores of 0.589 (diagnosis) and 0.405 (medication), representing 21.4% and 18.4% improvement over the best unimodal approach. Radiology reports consistently outperformed raw single-view chest radiographs as a supplementary modality. MLLMs demonstrated better zero- and few-shot performance than unimodal LLMs. Multi-view imaging consistently improved performance over single-view across all tasks. Conclusion The benefits of multimodal data integration are task-dependent. Healthcare LLMs should examine clinical data modalities according to specific tasks for efficient integration. These findings provide practical guidance for designing efficient clinical decision support systems.

Read PDF

Similar papers

#machine learning Preprint Sep 2026

Knowledge-Enriched Structured EHR Features for 30-Day Hospital Readmission Prediction on MIMIC-IV

Recent approaches to 30-day hospital readmission prediction rely on pre-trained language models applied to discharge summaries. Although these methods achieve strong performance, they depend on the availability of clinical notes, incur substantial computational costs, and yield representations that lack interpretabilit...

Mohamad Najafi, Hong-Yun Fu, M. Brochhausen et al. · 0 citations
Review Open access Oct 2024

Large Language Model Benchmarks in Medical Tasks

With the increasing application of large language models (LLMs) in the medical domain, evaluating these models' performance using benchmark datasets has become crucial. This paper presents a comprehensive survey of various benchmark datasets used in medical LLM tasks. These datasets span multiple modalities including t...

L. K. Yan, Qian Niu, Ming Li et al. · 32 citations · ⚡1
#artificial intelligence Preprint Sep 2026

MMTClinic: Multimodal, Multilingual Time Series Question Answering and Reasoning Benchmark for Clinical Domain

MMTClinic is presented, a benchmark designed to evaluate large language models (LLMs) on complex reasoning and question-answering tasks involving clinical time-series and reveals notable differences in model performance across tasks, languages, and modalities, highlighting current limitations in clinical reasoning capa...

Sourav Malakar, Harshit Nigam, Akash Ghosh et al. · 0 citations
Review Open access Sep 2026

Large language models as integrative intelligence for multimodal cardiovascular decision support

Background: In this narrative review, we examine large language models (LLMs) as an emerging component of cardiovascular artificial intelligence and propose the concept of integrative intelligence as a physician-supervised orchestration framework rather than a new model architecture. Current cardiovascular evidence rem...

P. Mitrović · 0 citations
Open access Aug 2026

Can GPT Be Used as an Alternative Prediction Model to Traditional Machine Learning and Neural Networks on Low-Volume Clinical Data?

The proposed GPT2-based table-to-text framework provides a practical and clinically interpretable approach for disease prediction from limited structured healthcare data and demonstrates strong potential for early risk detection, transparent clinical decision support, and reliable deployment in real-world low-resource...

S. Bin Akter, S. Akter, D. Eisenberg et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.