BreastGPT: A Multimodal Large Language Model for the Full Spectrum of Breast Cancer Clinical Routine
About
Breast cancer remains a leading cause of cancer-related mortality among women. Its clinical management requires multimodal reasoning across a clinical workflow that spans \textit{screening}, \textit{diagnosis} and \textit{treatment planning}, where each stage involves distinct imaging modalities, task objectives, and reasoning patterns. However, constrained by data scarcity and model versatility, existing medical MLLMs are typically evaluated on isolated modalities or narrow task families, limiting their ability to support workflow-level clinical reasoning. In this work, we first introduce \textbf{BreastStage}, a workflow-aligned breast imaging instruction corpus comprising 1.86M instruction-following pairs curated from 17 sub-datasets across 5 imaging modalities and 136 task templates. Its held-out split, \textbf{BreastStage-Bench}, provides a comprehensive benchmark for evaluating multimodal reasoning across the breast cancer care continuum. Building on this corpus, we propose \textbf{BreastGPT}, a unified MLLM equipped with a dual-branch visual encoder and concept-preserving token compression to bridge the scale gap between standard radiology and gigapixel pathology. On BreastStage-Bench, BreastGPT achieves 75.66\% closed-ended accuracy and 89.92\% open-ended score, outperforming both general-purpose and medical-specific MLLMs across clinical stages and task formats. These results suggest that workflow-aligned data and cross-scale visual modeling are critical for clinically grounded medical MLLMs. All data, code, and model checkpoints are released at https://yangyy-liu.github.io/BreastGPT.io.
Related benchmarks
| Task | Dataset | Result | Rank | |
|---|---|---|---|---|
| Grounded captioning | BreastStage (BUS) | Wtd79.32 | 32 | |
| Report Generation | BreastStage (MRI) | Wtd Score67.67 | 24 | |
| Visual Question Answering | BreastStage-Bench Closed-ended | Screening Accuracy (BUS)86.81 | 24 | |
| Visual Question Answering | BreastStage-Bench Open-ended | Screening (CT) Score95.29 | 24 | |
| Captioning | BreastStage (MRI) | Wtd Score67.69 | 23 | |
| Captioning | BreastStage CT | Wtd77.07 | 16 | |
| Captioning | BreastStage Histo | Wtd68.11 | 16 | |
| Captioning | BreastStage-Bench BUS | -- | 8 | |
| Captioning | BreastStage-Bench CT | -- | 8 | |
| Captioning | BreastStage-Bench Mammo | -- | 8 |