A Vision-Language Foundation Model for Precise and Comprehensive Brain Tumor Diagnosis from Preoperative Multimodal Data

📅 2026-09-14
📈 Citations: 0
Influential: 0
📄 PDF
🤖 AI Summary
研究开发了BrainVLM模型,利用MRI等多模态数据自动分类12种脑肿瘤,并提供诊断不确定性量化及报告生成,以改善术前非侵入性脑肿瘤诊断。
📝 Abstract
Background Non-invasive presurgical diagnosis of brain tumor types from Magnetic Resonance Imaging (MRI) is essential but challenging due to overlapping imaging features across tumor types, inter-observer variability, and the extensive training required for expertise. We aimed to develop an MRI-based Artificial Intelligence (AI) model for automatic and reliable brain tumor classification with diagnostic uncertainty quantification and radiology reports generation. Methods We developed BrainVLM to classify all 12 World Health Organization (WHO) 2021 brain tumor types. BrainVLM integrates an uncertainty quantification strategy to indicate prediction reliability and a module for generating radiology reports to elucidate the clinical rationale. BrainVLM was trained on multi-modal data (MRI scans, demographics, and radiology reports) from 40,043 individuals. It was validated on 5,211 patients with pathologically confirmed brain tumors, including 3,877 held-out patients from the primary hospital and 1,334 patients from 11 independent hospitals. We further conducted two proof-of-concept studies to validate its clinical utility in AI-clinician workflows: 1) a blinded multi-reader study where 12 neuroradiologists across varying experience levels interpreted 248 retrospective cases with or without AI assistance, and 2) a real-world prospective study in which 1,009 patients were independently and blindly assessed by BrainVLM and radiologists before surgery. Additionally, we demonstrated BrainVLM's utility in preoperative molecular subgroup prediction for adult-type diffuse gliomas, using a multi-center cohort of 632 patients.
Problem

Research questions and friction points this paper is trying to address.

Brain Tumor
MRI
Non-invasive Diagnosis
Inter-observer Variability
Training Expertise
Innovation

Methods, ideas, or system contributions that make the work stand out.

Vision-Language Foundation Model
Uncertainty Quantification
Radiology Reports Generation
Multi-modal Data Integration
Y
Yinong Wang
School of Computing and Data Science, The University of Hong Kong, China
J
Jianwen Chen
School of Computing and Data Science, The University of Hong Kong, China
Zhou Chen
Zhou Chen
Southeast University
Mechanism Design,Auction,Resource Allocation
S
Shuwen Kuang
Department of Oncology, Xiangya Hospital, Central South University, China
H
Haoning Jiang
School of Computing and Data Science, The University of Hong Kong, China
Yanzhao Shi
Yanzhao Shi
The University of Hong Kong
Multi-modal LLMMedical Report GenerationVision-Language Pretraining
H
Huichun Yuan
Changde Hospital, Xiangya School of Medicine, Central South University (The First People’s Hospital of Changde City), China
Y
Yan-ran Wang
Department of Biomedical Data Science, School of Medicine, Stanford University, Stanford, USA
B
Bing Wang
Department of Neurosurgery, The Second Affiliated Hospital, Hengyang Medical School, University of South China, China
Lei Wu
Lei Wu
University of Bristol, University College London
Functional soft materialsSoft roboticsSoft implantsBio-robots
B
Bin Tang
Department of Neurosurgery, The First Affiliated Hospital, Jiangxi Medical College, Nanchang University, China
L
Li Meng
Department of Radiology, Xiangya Hospital, Central South University, China
B
Baihua Luo
Department of Pathology, Xiangya Hospital, Central South University, China
B
Bin Zhou
Department of Neurosurgery, Xiangya Hospital, Central South University, Jiangxi (National Regional Center for Neurological Diseases), China
W
Wei Ding
The Affiliated Children’s Hospital of Xiangya School of Medicine, Hunan Children’s Hospital, China
W
Weiming Zhong
Department of Neurosurgery, Shenzhen Second People’s Hospital, China
W
Wei Hou
Department of Neurosurgery, First Hospital of Lanzhou University, China
Y
Yuanbing Chen
Department of Neurosurgery, The Third Xiangya Hospital, Central South University, China
Z
Zhiping Wan
Department of Neurosurgery, Tongji Hospital, School of Medicine, Tongji University, China
Wei Wang
Wei Wang
Tongji University
Image processing
Z
Zhenkun Xiao
Department of Neurosurgery, The Second Affiliated Hospital, Hengyang Medical School, University of South China, China
W
Wenwu Wan
Department of Neurosurgery, Chongqing Traditional Chinese Medicine Hospital, China
A
Allen He
Basis International School Park Lane Harbour, China
Yuyin Zhou
Yuyin Zhou
Assistant Professor, Computer Science and Engineering, Genomics Institute, UC Santa Cruz
medical image analysismachine learningcomputer visionAI in healthcare