Share your thoughts, 1 month free Claude Pro on usSee more
WorkDL logo mark

Health system learning achieves generalist neuroimaging models

About

Frontier artificial intelligence (AI) models, such as OpenAI's GPT-5 and Meta's DINOv3, have advanced rapidly through training on internet-scale public data, yet such systems lack access to private clinical data. Neuroimaging, in particular, is underrepresented in the public domain due to identifiable facial features within MRI and CT scans, fundamentally restricting model performance in clinical medicine. Here, we show that frontier models underperform on neuroimaging tasks and that learning directly from uncurated data generated during routine clinical care at health systems, a paradigm we call health system learning, yields high-performance, generalist neuroimaging models. We introduce NeuroVFM, a visual foundation model trained on 5.24 million clinical MRI and CT volumes using a scalable volumetric joint-embedding predictive architecture. NeuroVFM learns comprehensive representations of brain anatomy and pathology, achieving state-of-the-art performance across multiple clinical tasks, including radiologic diagnosis and report generation. The model exhibits emergent neuroanatomic understanding and interpretable visual grounding of diagnostic findings. When paired with open-source language models through lightweight visual instruction tuning, NeuroVFM generates radiology reports that surpass frontier models in accuracy, clinical triage, and expert preference. Through clinically grounded visual understanding, NeuroVFM reduces hallucinated findings and critical errors, offering safer clinical decision support. These results establish health system learning as a paradigm for building generalist medical AI and provide a scalable framework for clinical foundation models.

Akhil Kondepudi, Akshay Rao, Chenhui Zhao, Yiwei Lyu, Samir Harake, Soumyanil Banerjee, Rushikesh Joshi, Anna-Katharina Meissner, Renly Hou, Cheng Jiang, Asadur Chowdury, Ashok Srinivasan, Brian Athey, Vikas Gulani, Aditya Pandey, Honglak Lee, Todd Hollon• 2025

Related benchmarks

TaskDatasetResultRank
90-Day mRS PredictionICSPR-Stroke
Macro-Averaged AUROC0.714
4
Lesion-type classificationICSPR-Stroke
Macro AUROC90.2
4
Brain Age PredictionOpenBHB (test)
R20.673
4
Diagnosis and prognosisPublic 41 combs
AUROC74.1
4
Diagnosis and prognosisNYU Langone 30 combs
AUROC0.772
4
Diagnosis and prognosisNYU Long Island 30 combs
AUROC73.9
4
Diagnosis and prognosisBIND-MGH 45 combs
AUROC72.5
4
IDH mutation predictionUCSF-PDGM
Macro AUROC81.1
4
Multimodal LearningPublic datasets 12 tasks
AUROC0.748
4
Multimodal LearningBIND-MGH 30 tasks
AUROC74.2
4
Showing 10 of 12 rows

Other info

Follow for update