← Back to all news
arXiv Machine Learning August 24, 2026 By Jiazhuo Li, Yu Zhang, Yiming Fei, Kangkang Dong, Xiaojun Zhu, Houde Liu, Jinze Tao

Rethinking Demonstration Unlearning in Imitation Learning for Robotics

Read the original on arXiv Machine Learning →

The Flow has not summarised this story yet — read it at arXiv Machine Learning.

  • robotics

One email a morning, machine-written

One email a day, machine-written, one click to leave. We never share your address.

Related stories

arXiv Machine Learning
Jun 5

Auditing Demonstration Curation Metrics: Action-Only Scorers Fail on the Structural Defects That Degrade Imitation Policies

arXiv:2606. 05588v1 Announce Type: cross Abstract: Imitation-learning policies inherit the quality of the demonstrations they are trained on, and a growing set of curation metrics promise to score and filter low-quality demonstrations automatically.

By Aarav Bedi (University of California, Berkeley)
More like this →
arXiv Machine Learning
Aug 6

Suppression Sticks, Locality Is Fragile: A Closed-Loop Target-and-Control Audit of Task-Vector Negation in VLA Policies

arXiv:2608. 04692v1 Announce Type: cross Abstract: Task-vector arithmetic offers a closed-form way to modify a model, yet its behavioral locality remains unclear in closed-loop robot control.

By Shaoguang Wang, Weiyu Guo, Rushi Dai, Yiren Zhao, Yandong Guo, Hui Xiong
roboticsmultimodal
More like this →
arXiv AI
Sep 10

DISEIL: Demonstration Distillation for Sample-Efficient Imitation Learning

arXiv:2609.08123v1 Announce Type: cross Abstract: A robot that can be taught a new task from a handful of demonstrations has to work out for itself what it still cannot do, and then ask for exactly t...

By Suyog Khanal, Arun Kumar A V, Santu Rana
llmsroboticsefficiencymultimodal
More like this →
arXiv AI
Aug 10

AutoIntervene: Calibrated Intervention for Action-Chunking Imitation Learning Policies

arXiv:2608. 07065v1 Announce Type: cross Abstract: Action-chunking visuomotor policies learn from demonstrations and improve temporal consistency by predicting short action sequences rather than single-step commands.

By Jinhe Tang, Weiming Zhi
ragrobotics
More like this →
arXiv AI
Jun 30

Behavior Uncloning: Distilling Mode Redirection into Policy Weights without Inference-Time Steering

arXiv:2606. 29201v1 Announce Type: cross Abstract: Behavior-cloned policies often learn multiple behavior modes from demonstration datasets, including modes that are unsafe or otherwise undesired at deployment.

By Hao Wang, Jiuzhou Lei, Dayou Li, Bangya Liu, Minghui Zheng, Manling Li, Ruohan Zhang, Zhiwen Fan
diffusionroboticsefficiency
More like this →
arXiv Machine Learning
Jun 3

How Visible Are Silent Manipulation Failures? An Observability Study of False-Success Detection in Simulated Robot Episodes

arXiv:2606. 03134v1 Announce Type: cross Abstract: Imitation-learning policies for robot manipulation inherit the quality of the success labels attached to their training episodes, and those labels are usually produced by the robot's own success check.

By Aarav Bedi (University of California, Berkeley)
robotics
More like this →
About Pricing API Newsletter Sources Privacy Terms Refunds Accessibility Provider info Contact RSS

The Flow links to publishers and never republishes their articles. Summaries are machine-generated.

v1.1.0 · 5f852ea