Hugging Face Trending Papers

ProtoPointNet: Prototype-Based Interpretable Classification of 3D Dental Point Clouds with Verifiable Spatial Activations

Prototype-based networks provide inherently interpretable classification by linking predictions to learned exemplars, but their use in 3D point clouds and clinical surface-pair reasoning remains limited. We introduce ProtoPointNet, a prototype-based model for dental occlusion classification from registered upper--lower intraoral arch pairs.

arXiv AI
Sep 17

Automated Dental Caries Segmentation in Panoramic Radiographs Using Dual-Stage Deep Learning

The paper introduces a dual‑stage deep learning system for detecting dental caries in panoramic radiographs. It first localizes teeth using Faster R‑CNN, then applies U‑Net for pixel‑wise caries segmentation, converting polygon annotations into high‑resolution binary masks. Trained on 3,000 images with both expert and algorithmic labels, the model achieves an IoU of 0.9013, Dice of 0.9482, Recall of 0.9433, and Precision of 0.9774, outperforming existing methods and reducing false positives.

By Jihun Kim, Kyeonghun Kim, Jong-yeol Lee, Yeongseok Seo, Dohyun Chun
arXiv Computer Vision
Aug 21

PhysSFI-Net: Physics-informed Geometric Learning of Skeletal and Facial Interactions for Orthognathic Surgical Outcome Prediction

arXiv:2601. 02088v3 Announce Type: replace Abstract: Orthognathic surgery repositions jaw bones to restore occlusion and enhance facial aesthetics.

By Jiahao Bao, Huazhen Liu, Yu Zhuang, Leran Tao, Xinyu Xu, Yongtao Shi, Mengjia Cheng, Yiming Wang, Congshuang Ku, Ting Zeng, Yilang Du, Siyi Chen, Shunyao Shen, Suncheng Xiang, Hongbo Yu
arXiv Computer Vision
Sep 4

TokenMatch: 3D Mesh Correspondence Transformer with Curvature-Guided Tokenisation

TokenMatch is a transformer-based model that estimates 3D shape correspondences by adaptively tokenising meshes into curvature-guided patches. Trained only on the BeCoS partial-to-partial dataset, it generalises to full-shape matching without retraining, using self‑ and cross‑attention to learn patch‑ and point‑level relations. Evaluated on CP2P, PSMAL, BeCoS, FAUST, SCAPE, and SHREC'19, TokenMatch consistently outperforms existing methods in mean geodesic error and intersection‑over‑union while achieving sub‑second inference speeds.

By Adeela Islam, Zorah L\"ahner, Vittorio Murino, Vladislav Golyanik
arXiv Computer Vision
Sep 17

AgenTeeth: A Model-Agnostic Framework for Suppressing Hallucination in Frozen Vision-Language Models on Dental X-Rays via Tool Evidence Injection

arXiv:2609.17800v1 Announce Type: new Abstract: Vision-language models (VLMs) remain largely unreliable on panoramic dental radiographs and can rely on learned anatomical priors rather than evidence...

By Ahmed Rafid, Fariya Ahmed, Rumman Adib, Mehedi Ahamed, Ajwad Abrar, Tareque Mohmud Chowdhury