Teeth-SEG: An Efficient Instance Segmentation Framework for Orthodontic Treatment Based on Multi-Scale Aggregation and Anthropic Prior Knowledge
Bo Zou, Shaofeng Wang, Hao Liu, Gaoyue Sun, Yajie Wang, Feifei Zuo, Chengbin Quan, Youjian Zhao
Abstract
Teeth localization, segmentation, and labeling in 2D images have great potential in modern dentistry to enhance dental diagnostics, treatment planning, and population-based studies on oral health. However, general instance segmentation frameworks are incompetent due to 1) the sub-tle differences between some teeth’ shapes (e.g., maxillary first premolar and second premolar), 2) the teeth's position and shape variation across subjects, and 3) the presence of abnormalities in the dentition (e.g., caries and edentulism). To address these problems, we propose a ViT-based frame-work named TeethSEG, which consists of stacked Multi-Scale Aggregation (MSA) blocks and an Anthropic Prior Knowledge (APK) layer. Specifically, to compose the two modules, we design a unique permutation-based upscaler to ensure high efficiency while establishing clear segmentation boundaries with multi-head selflcross-gating layers to emphasize particular semantics meanwhile maintaining the divergence between token embeddings. Besides, we collect the first open-sourced intraoral image dataset IO150K, which comprises over 150k intraoral photos, and all photos are annotated by orthodontists using a human-machine hybrid algorithm. Experiments on IO150K demonstrate that our TeethSEG outperforms the state-of-the-art segmentation models on dental image segmentation.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 92a09fb6-0ac1-4b23-b7e0-dc4e8e2e2dfbBuilds on9
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- SegFormer: Simple and Efficient Design for Semantic Segmentation with TransformersEnze Xie, Wenhai Wang, Zhiding Yu, Anima Anandkumar et al.NeurIPS 2021 · 9,661 citations
- Vision Transformers for Dense PredictionRené Ranftl, Alexey Bochkovskiy, Vladlen KoltunICCV 2021 · 2,647 citations
- Vision Transformer Adapter for Dense PredictionsZhe Chen, Yuchen Duan, Wenhai Wang, Junjun He et al.ICLR 2023 · 204 citations
- DArch: Dental Arch Prior-assisted 3D Tooth Instance Segmentation with Weak AnnotationsLiangdong Qiu, Chongjie Ye, Pei Chen, Yunbi Liu et al.CVPR 2022 · 34 citations
Related papers
- 3DTeethSAM: Taming SAM2 for 3D Teeth SegmentationZhiguo Lu, Jianwen Lou, Mingjun Ma, Hairong Jin et al.AAAI 2026 · 1 citation
- Photo-Guided Tooth Segmentation on 3D Oral Scan ModelShaojie Zhuang, Guangshun Wei, Jiangxin He, Yuanfeng ZhouCVPR 2026
- TeethGenerator: A Two-Stage Framework for Paired Pre- and Post-Orthodontic 3D Dental Data GenerationChangsong Lei, Yaqian Liang, Shaofeng Wang, Jiajia Dai et al.ICCV 2025 · 1 citation
- TSGCNet: Discriminative Geometric Feature Learning With Two-Stream Graph Convolutional Network for 3D Dental Model SegmentationLingming Zhang, Yue Zhao, Deyu Meng, Zhiming Cui et al.CVPR 2021
- 3D Dental Model Segmentation with Geometrical Boundary PreservingShufan Xi, Zexian Liu, Junlin Chang, Hongyu Wu et al.CVPR 2025
