Native Intelligence Emerges from Large-Scale Clinical Practice: A Retinal Foundation Model with Deployment Efficiency

📅 2025-12-16
📈 Citations: 0
Influential: 0
📄 PDF
🤖 AI Summary
Existing retinal foundation models rely on manually annotated, curated datasets and require extensive task-specific fine-tuning, hindering deployment in resource-constrained clinical settings. This paper introduces ReVision—the first retinal foundation model trained exclusively on real-world, decade-long telemedicine data (486,000 fundus color photographs with corresponding clinical reports), eliminating the need for manual annotations and enabling zero-shot disease detection and cross-institutional, cross-modal generalization. We propose the “clinically native intelligence” paradigm, replacing curated datasets with authentic consultation data, and integrate contrastive vision-language alignment, zero-shot prompting, lightweight adapter-based fine-tuning, and cross-domain representation transfer. ReVision achieves a zero-shot AUROC of 0.946 across 12 public benchmarks and 0.952 on three independent clinical cohorts; it improves physician diagnostic accuracy by 14.8%; and matches full fine-tuning performance using only a minimal number of trainable parameters and scarce annotations.

Technology Category

Application Category

📝 Abstract
Current retinal foundation models remain constrained by curated research datasets that lack authentic clinical context, and require extensive task-specific optimization for each application, limiting their deployment efficiency in low-resource settings. Here, we show that these barriers can be overcome by building clinical native intelligence directly from real-world medical practice. Our key insight is that large-scale telemedicine programs, where expert centers provide remote consultations across distributed facilities, represent a natural reservoir for learning clinical image interpretation. We present ReVision, a retinal foundation model that learns from the natural alignment between 485,980 color fundus photographs and their corresponding diagnostic reports, accumulated through a decade-long telemedicine program spanning 162 medical institutions across China. Through extensive evaluation across 27 ophthalmic benchmarks, we demonstrate that ReVison enables deployment efficiency with minimal local resources. Without any task-specific training, ReVision achieves zero-shot disease detection with an average AUROC of 0.946 across 12 public benchmarks and 0.952 on 3 independent clinical cohorts. When minimal adaptation is feasible, ReVision matches extensively fine-tuned alternatives while requiring orders of magnitude fewer trainable parameters and labeled examples. The learned representations also transfer effectively to new clinical sites, imaging domains, imaging modalities, and systemic health prediction tasks. In a prospective reader study with 33 ophthalmologists, ReVision's zero-shot assistance improved diagnostic accuracy by 14.8% across all experience levels. These results demonstrate that clinical native intelligence can be directly extracted from clinical archives without any further annotation to build medical AI systems suited to various low-resource settings.
Problem

Research questions and friction points this paper is trying to address.

Develops retinal AI using real clinical data instead of curated datasets
Enables efficient deployment in low-resource medical settings
Creates foundation model for multiple ophthalmic tasks without retraining
Innovation

Methods, ideas, or system contributions that make the work stand out.

Leveraging large-scale telemedicine data for training
Achieving zero-shot disease detection without task-specific training
Enabling efficient deployment with minimal local adaptation
🔎 Similar Papers
No similar papers found.
J
Jia Guo
School of Biomedical Engineering, Tsinghua Medicine, Tsinghua University, Beijing, China
Jiawei Du
Jiawei Du
National Taiwan University; ex-Intern @ Samsung Research
Speech processingNeural codingGenerative AIAI security
S
Shengzhu Yang
School of Medical Technology, Beijing Institute of Technology, Beijing, China
S
Shuai Lu
School of Information and Electronics, Beijing Institute of Technology, Beijing, China
W
Wenquan Cheng
School of Biomedical Engineering, Tsinghua Medicine, Tsinghua University, Beijing, China
Kaiwen Zhang
Kaiwen Zhang
Associate Professor, Software/IT Engineering, École de technologie supérieure
blockchainsdistributed systemspublish/subscribeonline gamesmiddleware
Y
Yihua Sun
School of Biomedical Engineering, Tsinghua Medicine, Tsinghua University, Beijing, China
C
Chuhong Yang
School of Information and Electronics, Beijing Institute of Technology, Beijing, China
Weihang Zhang
Weihang Zhang
Assistant Professor, School of Medical Technology, Beijing Institute of Technology
medical image processing
F
Fang Chen
School of Biomedical Engineering, Shanghai Jiaotong University, Shanghai, China
Y
Yilan Wu
Beijing Visual Science and Translational Eye Research Institute (BERI), Beijing Tsinghua Changgung Hospital Eye Center, School of Clinical Medicine, Tsinghua Medicine, Tsinghua University
Lie Ju
Lie Ju
University College London; Moorfields Eye Hospital; Monash University
Computer VisionMedical Image AnalysisOphthalmology
Guochen Ning
Guochen Ning
Tsinghua University
medical robotAutonomous System
L
Longfei Ma
School of Biomedical Engineering, Tsinghua Medicine, Tsinghua University, Beijing, China
H
Huiping Yao
Department of Ophthalmology, Ruijin Hospital, Shanghai Jiao Tong University School of Medicine, Shanghai, China
J
Jinyuan Wang
Beijing Visual Science and Translational Eye Research Institute (BERI), Beijing Tsinghua Changgung Hospital Eye Center, School of Clinical Medicine, Tsinghua Medicine, Tsinghua University
P
Peilun Shi
Department of Biomedical Engineering, The Chinese University of Hong Kong, Hong Kong SAR, China
Y
Yukun Zhou
Institute of Ophthalmology, University College London, London, UK
J
Jie Xu
Beijing Tongren Hospital, Capital Medical University, Beijing, China
P
Pearse A. Keane
Institute of Ophthalmology, University College London, London, UK
H
Hanruo Liu
Henan Provincial People’s Hospital, Henan Eye Hospital, Henan, China
H
Hongen Liao
School of Biomedical Engineering, Tsinghua Medicine, Tsinghua University, Beijing, China
Ningli Wang
Ningli Wang
Beijing Tongren Hospital & Henan Academy of Innovations in Medical Science (AIMS)
GlaucomaClinical OphthalmologyEye DiseasesPublic HealthDigital Health
H
Huiqi Li
School of Information and Electronics, Beijing Institute of Technology, Beijing, China