arXiv Computer Vision By Yitong Xing, Yuhao Cheng, Yanping Li, Yichao Yan

Enhanced Knowledge Distillation for Detection Transformer via Teacher Prediction Refinement

Read the original on arXiv Computer Vision →

The paper introduces Teacher Prediction Refinement Distillation (TPRD), a plug‑in module for Detection Transformers that refines teacher predictions before distillation. TPRD corrects degraded positive predictions and suppresses overconfident negatives, while preserving informative dark knowledge through Maximum Dark Knowledge Preservation. Experiments on MS COCO and PASCAL VOC show that these refinements improve the quality of supervision and the resulting student model’s performance.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Computer Vision.

arXiv AI
Jul 28

Rethinking Classifier-Free Guidance in On-Policy Diffusion Distillation

arXiv:2607. 24731v1 Announce Type: cross Abstract: On-policy distillation (OPD) adapts diffusion models by querying a teacher along trajectories generated by the current student, but how it should behave under classifier-free guidance (CFG), a default component of modern diffusion systems, remains poorly understood.

By Bingnan Li, Haozhe Wang, Haozhong Xiong, Fangtai Wu, Jinpeng Yu, Yang Shi, Jiaming Liu, Ruihua Huang
arXiv Computer Vision
Sep 3

CA-OPD: Confidence-Aware On-Policy Distillation for Structured Visual Prediction

CA-OPD is a confidence‑aware on‑policy distillation framework that improves structured visual prediction by using teacher confidence to selectively correct unreliable student transitions and gradually transfer rollout control to the student. The method aligns supervision with intervention decisions, providing direct cross‑entropy loss for corrected tokens and full predictive distribution for retained tokens. In a multi‑teacher setting for GUI grounding and OCR, CA‑OPD significantly outperforms the Qwen3.5‑0.8B baseline, achieving large gains on benchmarks such as ScreenSpot‑Pro and OCRBench‑v2 English.

By Menghao Li, Linjie Mu, Yin Wang, Haotian Hu, Yannian Gu, Lujiayi Xue, Fanyi Wang