ArchMap: Arch-Flattening and Knowledge-Guided Vision Language Model for Tooth Counting and Structured Dental Understanding

📅 2025-11-18
📈 Citations: 0
Influential: 0
📄 PDF
🤖 AI Summary
Semantic parsing of intraoral 3D scans is hindered by scarce annotated data, device heterogeneity, geometric incompleteness, and absent texture. Method: We propose a training-free structured understanding framework that integrates a geometry-aware dental-arch flattening module to generate multi-view projections, and constructs a Dental Knowledge Base (DKB) unifying tooth-level ontologies, eruption-stage rules, and clinical semantics. Crucially, we pioneer ontology-guided visual-language modeling by embedding domain-specific ontological knowledge into a vision-language model (VLM) to enable knowledge-constrained reasoning. Results: Evaluated on 1,060 orthodontic cases, our method significantly outperforms supervised models and prompt-based VLM baselines across teeth counting, anatomical segmentation, eruption staging, and caries/edentulism identification—particularly under low-quality scans—demonstrating superior robustness and cross-device generalizability.

Technology Category

Application Category

📝 Abstract
A structured understanding of intraoral 3D scans is essential for digital orthodontics. However, existing deep-learning approaches rely heavily on modality-specific training, large annotated datasets, and controlled scanning conditions, which limit generalization across devices and hinder deployment in real clinical workflows. Moreover, raw intraoral meshes exhibit substantial variation in arch pose, incomplete geometry caused by occlusion or tooth contact, and a lack of texture cues, making unified semantic interpretation highly challenging. To address these limitations, we propose ArchMap, a training-free and knowledge-guided framework for robust structured dental understanding. ArchMap first introduces a geometry-aware arch-flattening module that standardizes raw 3D meshes into spatially aligned, continuity-preserving multi-view projections. We then construct a Dental Knowledge Base (DKB) encoding hierarchical tooth ontology, dentition-stage policies, and clinical semantics to constrain the symbolic reasoning space. We validate ArchMap on 1060 pre-/post-orthodontic cases, demonstrating robust performance in tooth counting, anatomical partitioning, dentition-stage classification, and the identification of clinical conditions such as crowding, missing teeth, prosthetics, and caries. Compared with supervised pipelines and prompted VLM baselines, ArchMap achieves higher accuracy, reduced semantic drift, and superior stability under sparse or artifact-prone conditions. As a fully training-free system, ArchMap demonstrates that combining geometric normalization with ontology-guided multimodal reasoning offers a practical and scalable solution for the structured analysis of 3D intraoral scans in modern digital orthodontics.
Problem

Research questions and friction points this paper is trying to address.

Standardizing raw 3D dental meshes with varying arch poses and incomplete geometry
Eliminating dependency on large annotated datasets and modality-specific training
Enabling robust structured dental analysis across different devices and clinical conditions
Innovation

Methods, ideas, or system contributions that make the work stand out.

Geometry-aware arch-flattening module standardizes 3D meshes
Dental Knowledge Base encodes hierarchical tooth ontology
Training-free framework combines geometric normalization with multimodal reasoning
🔎 Similar Papers
No similar papers found.
B
Bohan Zhang
School of AI and Advanced Computing, Xi’an Jiaotong-Liverpool University, China
Y
Yiyi Miao
School of AI and Advanced Computing, Xi’an Jiaotong-Liverpool University, China; School of Electrical Engineering, Electronics and Computer Science, University of Liverpool, United Kingdom
T
Taoyu Wu
School of Advanced Technology, Xi’an Jiaotong-Liverpool University, China; School of Physical Sciences, University of Liverpool, Liverpool, United Kingdom
T
Tong Chen
School of AI and Advanced Computing, Xi’an Jiaotong-Liverpool University, China; School of Electrical Engineering, Electronics and Computer Science, University of Liverpool, United Kingdom
Ji Jiang
Ji Jiang
School of Mathematics and Physics, Xi’an Jiaotong-Liverpool University, China
Z
Zhuoxiao Li
Urban Governance and Design Thrust, The Hong Kong University of Science and Technology (Guangzhou), China
Zhe Tang
Zhe Tang
University of Liverpool
WSNIoThybrid network
Limin Yu
Limin Yu
Xi'an Jiaotong-Liverpool University
sonar detectionrational waveletsmedical image analysisAGV system design
Jionglong Su
Jionglong Su
Xi'an Jiaotong-Liverpool University
AI Big Data Machine Learning Statistics