arXiv AI By Vicky Feliren, A. Taufiq Asyhari, Muhamad Risqi U. Saputra

ENCP: Episode-Normalized Conformal Prediction for Vision-and-Language Navigation

Read the original on arXiv AI →

The Flow has not summarised this story yet — read it at arXiv AI.

arXiv Computation and Language
Sep 2

IntroConformal: Conformal Factuality Guarantees for Large Vision-Language Models via Introspective Signals

IntroConformal introduces a training‑free Conformal Risk Control framework that offers finite‑sample, distribution‑free factuality guarantees for Large Vision‑Language Models. It uses introspective signals—layer‑wise semantic stability and verification probability derived from the model’s own hidden states—to assess claim factuality. Experiments across multiple LVLM architectures show that IntroConformal meets the conformal risk guarantee while reducing abstention and matching or surpassing external verifier baselines in claim‑level discrimination.

By Md. Atabuzzaman, Christian Alexander, Chris Thomas
arXiv Computer Vision
Sep 22

What do VLM-Based Vision-Language Navigation Models Rely on: Interpreting and Steering Policy Behavior

arXiv:2609.24576v1 Announce Type: cross Abstract: Modern Vision-Language Navigation (VLN) models rely mostly on pre-trained large Vision-Language Models (VLMs) to predict navigation actions. While th...

By D\'ebora Oliveira Makowski, Samiran Gode, Abhijeet Nayak, Marco Hutter, Cordelia Schmid, Lukas Rosenberger Schmid, Wolfram Burgard
arXiv Computer Vision
Sep 18

Latent-Centroid Steering: Single-Pass Classifier-Free Guidance for Command-Aligned Autonomous Driving

The paper introduces Latent-Centroid Steering (LCS), a single-pass classifier-free guidance method for vision‑language autonomous driving models. LCS replaces instance‑level residuals with class‑level latent shifts, projecting conditional representations toward precomputed command‑specific centroids to enhance command adherence. Experiments on Bench2Drive and nuScenes show that LCS cuts inference latency by about 50% while improving driving performance.

By Meibo Hu, Jiamian Wang, Pichao Wang, Zhiqiang Tao