arXiv Computer Vision
Aug 31

Physics-Guided Flow Matching for CT Image Reconstruction

The paper introduces a high‑resolution Rectified Flow Matching model trained on 256×256 chest CT images to serve as a generative prior for CT reconstruction. A two‑stage training strategy—initial strong anatomically informed augmentation followed by fine‑tuning—helps mitigate overfitting and improve structural fidelity. When evaluated on various CT inverse problems, Flow Matching‑based reconstruction methods outperform diffusion‑based algorithms in PSNR, SSIM, and perceptual quality while requiring fewer sampling steps.

By Davide Evangelista
arXiv Computer Vision
Sep 11

GRADE: Single-Frame Generative Radar Depth Estimation Under Visual Degradation

GRADE is a method for estimating high‑fidelity metric depth from a single radar frame, even when visual sensors fail due to smoke, fog, or darkness. It first converts raw 4D radar spectra into coarse depth, then uses a latent diffusion model conditioned on this estimate to recover fine structural detail. A pixel‑space adapter incorporates any available camera cues and is trained across clear, smoke‑degraded, and occluded inputs, allowing the output to rely more on radar as visibility worsens. On a dataset of ~95K frames from 12 buildings with real smoke, GRADE achieves an MAE of 0.303 m in clear scenes and 0.313 m under smoke, outperforming existing baselines.

By Bin Zhao, Patrick Chiou, Nakul Garg