arXiv Machine Learning

Scalable and Differentiable Point-Cloud Registration Using Maximum Mean Discrepancy

arXiv:2606. 27818v1 Announce Type: cross Abstract: We present MMD-Reg, a novel correspondence-free approach to point-cloud registration that is differentiable and has linear computational complexity in the number of points.

arXiv Computer Vision
Aug 28

Test Time Adaptation Methods for Point Cloud Registration in Laparoscopic Surgery

The paper investigates test‑time adaptation (TTA) techniques for 3D point‑cloud registration in laparoscopic surgery, where synthetic training data must be adapted to noisy, sparse, and occluded real intraoperative reconstructions. It adapts three families of TTA methods—model, normalization, and input adaptation—to handle asymmetric shifts between preoperative meshes and intraoperative clouds, replacing classification‑based entropy objectives with correspondence‑based ones. Experiments on synthetic and real targets show that input adaptation consistently reduces registration error with low inference latency, making it the most promising approach for surgical applications.

By Nina Bodelot, Soufiane Belharbi, Eric Granger
arXiv AI
Sep 17

Mask 2D-3D: Adaptive Dual-Masked Autoencoder Network for Image-to-Point Cloud Registration

The paper introduces Mask 2D-3D, an Adaptive Dual-Masked Autoencoder Network designed for image-to-point cloud registration. It proposes an Intermodal Dual-MAE Framework (ID-MAE) with a Similarity-based RL Masking Strategy (SRLM) that adaptively masks informative positions using cross-modal similarity and reinforcement learning. Experiments on RGB-D Scenes v2 and 7-Scenes benchmarks demonstrate state-of-the-art performance in this registration task.

By Zhixin Cheng, Jiacheng Deng, Xiaotian Yin, Baoqun Yin, Richang Hong, Tianzhu Zhang
Hugging Face Trending Papers
Aug 3

An Accessible Solution for Deformable Image Registration Compared with Learning-Based Approaches

Deformable image registration (DIR) is a core problem in medical image analysis; but, unlike labeling decision problems such as classification and segmentation, registration is a problem class that involves stringent physical constraints. Although deep learning methods have made faster registration possible, the resulting models are often difficult to interpret compared to hand-crafted methods with explicit objectives and interpretable physical meaning.

arXiv Computer Vision
Sep 14

Spectral Consistency-Guided Multiview Point Cloud Registration for Low-Overlap Scenes

The paper introduces GMPCR, a non‑learning spectral consistency‑guided framework for multiview point cloud registration in low‑overlap scenes. GMPCR refines initial correspondences into a second‑order compatibility structure, uses spectral analysis to filter unreliable matches and select informative scan pairs, and then applies maximal‑clique hypothesis generation for robust relative transformations. The resulting sparse pose graph is further refined with an adaptive history‑aware synchronization scheme, and a recovery mechanism allows previously down‑weighted edges to regain confidence, achieving high registration recalls on benchmark datasets while reducing computational cost.

By Tianyu Li, Yanghong Lin, Shudong Zhou, Kui Yang, Jingru Zhang, Li Fang, Wei Yao
arXiv Computer Vision
Sep 18

BINDER: A Latent Variable Model for Probabilistic Medical Image Registration

BINDER is a new probabilistic model for medical image registration that builds on mutual information and uses latent voxel‑wise correspondences to enable closed‑form iterative updates. The approach yields a demons‑like optimization algorithm that performs robustly on both monomodal and multimodal tasks, and a sampler that quantifies uncertainty in high‑dimensional 3D deformations. The authors provide the code on GitHub for public use.

By Stefano Cerri, Amirhossein Hassankhani, Ya\"el Balbastre, Koen Van Leemput
arXiv Computer Vision
Sep 18

UniReg: Conditional Unified Model for Medical Image Registration

UniReg is a conditional unified model for medical image registration that adapts deformation field estimation based on anatomical priors, registration type constraints, and instance-specific features. It combines the precision of task‑specific learning with the generalization of traditional optimization, enabling effective alignment across diverse CT and MR scenarios within a single framework. Experiments show UniReg outperforms state‑of‑the‑art learning‑based methods in accuracy while providing strong cross‑scenario generalization and reducing training cost and model redundancy.

By Zi Li, Jianpeng Zhang, Tai Ma, Tony C. W. Mok, Yan-Jie Zhou, Zeli Chen, Xianghua Ye, Le Lu, Cheng Chen, Dakai Jin
arXiv Computer Vision
Sep 24

DMM-Align: Closed-Loop Optimization for 2D-3D Registration with Dual-Role Diffusion

DMM-Align introduces a closed‑loop framework for 2D‑3D registration that jointly refines correspondences, estimates pose, and learns representations using a shared differentiable geometric state. The method employs two diffusion processes: a geometry‑aware diffusion that improves the soft matching matrix for robust correspondence estimation, and a geometry‑conditioned diffusion teacher that feeds pose‑induced supervision back into feature learning. Experiments on 7‑Scenes and RGB‑D Scenes V2 show that DMM‑Align outperforms strong baselines, particularly in low‑overlap and heavily occluded scenarios, demonstrating the value of closed‑loop geometric feedback.

By Chongjian Wang, Junjie Gao