arXiv:2609.34182v2 Announce Type: replace-cross
Abstract: Dexterous manipulation requires tactile feedback. However, robot tactile demonstrations are difficult to scale,because dexterous-hand teleope...
By Wenqiao Li, Qianyou Zhao, Jiawen Hao, Xuezhou Zhu, Tengyu Liu, Kaifeng Zhang, Chuan Wen, Siyuan Huang
arXiv:2609.36421v1 Announce Type: cross
Abstract: Reinforcement learning (RL) is a powerful framework for robotic control, yet its practical application is often hindered by high sample complexity. T...
By Rayan Mazouz, Haibo Zhao, Chris Hillar, Christian Shewmake
arXiv:2609.36540v1 Announce Type: cross
Abstract: Generalist robot policies such as vision-language-action models (VLAs) have achieved remarkable generalization, but their inference delays can confli...
By Moritz Zoellner, Reece O'Mahoney, Ioannis Havoutis, Rohan Paleja
arXiv:2609.36145v1 Announce Type: new
Abstract: Tampered Text Detection (TTD) is essential for safeguarding document authenticity in security-critical workflows. Existing expert models are effective...
By Kaiqing Lin, Songze Li, Shen Chen, Yunfei Guo, Xiaoye Qiu, Haodong Li, Taiping Yao, Bo Wang, Youchang Xiao, Bin Li, Shouhong Ding
arXiv:2609.36875v1 Announce Type: new
Abstract: Accurate instance segmentation in dynamic scenes is important for downstream applications such as robotics and autonomous driving. Existing Segment Any...
By Jingdong Zhang, Xin Li, Jan Kautz, Wenping Wang, Chris Choy
arXiv:2609.36851v1 Announce Type: new
Abstract: End-to-end autonomous driving policies are commonly trained via imitation learning on logged demonstrations without observing the consequences of their...
By Hongbin Lin, Chaoda Zheng, Yiming Yang, Xiangyu Li, Shijia Chen, Jinhao Deng, Kangjie Chen, Dongbin Zhang, Jie Feng, Yu Zhang, Xianming Liu, Shuguang Cui, Boyang Wang, Zhen Li
arXiv:2609.36906v1 Announce Type: new
Abstract: Reliable embodied decisions under partial observability require informative observations and sufficient supporting evidence. However, semantic scores a...
By Sean Hardesty Lewis, Zuyi Guo, Benwang Chen, Zirui Liu, Hongyi Lin, Heye Huang
arXiv:2609.36940v1 Announce Type: new
Abstract: Accurate dynamic scene reconstruction is important for robotic perception, where temporally consistent representations of dynamic environments are esse...
By Thai Duy Nguyen, Haitian Zhang, Addison Lin Wang
arXiv:2607.28225v2 Announce Type: replace
Abstract: Agentic vision-language models (VLMs), which interleave textual reasoning with explicit tool calls such as cropping and code-based image manipulati...
By Haoqing Wang, Xingrun Xing, Ziheng Li, Jianyuan Guo, Yehui Tang
arXiv:2606.08495v2 Announce Type: replace-cross
Abstract: Humanoid robots require whole-body motions that adapt to scene context, task requirements, and user intent. Motion tracking reproduces specif...
By Haoyang Ge, Peng Ren, Yukun Shi, Cong Huang, Kun Li, Kai Chen
The paper introduces Embodied Semantic Communication (ESC), a new paradigm that redefines information transmission for autonomous agents by embedding multimodal perceptual states, hardware capabilities, and collaborative intents into unified, action‑oriented semantic representations. ESC enables heterogeneous agents to parse, align, and ground shared information directly into local motor control, addressing the limitations of traditional communication approaches that focus solely on bit delivery or single‑task optimization. The tutorial outlines ESC’s conceptual boundaries, system characteristics, and technical pathways, mapping relevant mathematical tools such as semantic information theory, world models, and multi‑agent decision theory, and concludes with a roadmap of open challenges like semantic reliability, dynamic interaction, and bandwidth‑adaptive transmission.
By Yizheng Huang, Wensheng Lin, Lixin Li, Qinghe Du, Wenchi Cheng, Zhu Han
arXiv:2607.28623v2 Announce Type: replace-cross
Abstract: We present PAC-MAN, a perception-aware CBF-RL framework that couples control-barrier safety with deployment-realistic onboard sensing for who...
By Lizhi Yang, Junheng Li, Aaron D. Ames
arXiv:2609.32453v2 Announce Type: replace-cross
Abstract: Robotic manipulation is inherently history-dependent, yet most pretrained robotic policies condition on only the current observation or a sho...
By Xinyu Zhao, Yixiang Shan, Tao Yang, Runyu Lei, Yiming Zhao, Jiaxin Fan, Zongbao Feng, Peng Jia
arXiv:2609.35955v1 Announce Type: new
Abstract: Understanding human-entity interactions requires recovering each person-action event's participants, roles, and shared identities. This structure can s...
By Di Wen, Wenhao Guo, Yuedong Tan, Yun Huang, Minheng Wu, Zhihang Chen, Haiwen Sun, Fei Teng, Zhiyuan Gao, Yufeng Zhang, Yuanhao Luo, Jingqi Zhang, Yufan Chen, Junwei Zheng, Ruiping Liu, Jiale Wei, Kailun Yang, Kunyu Peng
arXiv:2401.05018v3 Announce Type: replace
Abstract: Human motion prediction is a crucial capability for advanced robotic systems that interact with humans. In facilities with dynamic human-robot coll...
By Sarmad Idrees, Seokman Sohn, Jongeun Choi
arXiv:2609.38016v1 Announce Type: new
Abstract: Constrained Reinforcement Learning has recently gained increasing attention in the field of Safe Autonomous Driving, where the general mechanism is to...
By Huan Rong, Chao Yin, Anouar Imel, Yijie Xia, Tinghuai Ma
arXiv:2609.36518v1 Announce Type: cross
Abstract: Robots must often continue a task after a target moves, the viewpoint shifts, or an obstacle appears, even though their earlier observations and comm...
By Yunbei Zhang, Zijian Jin, Yuanzhe Liu, Janet Wang, Xilun Zhang, Yuyou Zhang, Zhenyu Zhang, Daoan Zhang, Shuaicheng Niu, Gen Li, Jianfei Yang, Jihun Hamm, Ismini Lourentzou, Weirui Ye, Bo Liu, Peter Stone, Marco Pavone
arXiv:2609.37011v1 Announce Type: cross
Abstract: Federated learning enables multiple clients to collaboratively train models without sharing their private data. However, the lack of visibility into...
By Hongxu Su, Jianzhu Yao, Xuechao Wang, Pramod Viswanath
arXiv:2609.37264v1 Announce Type: cross
Abstract: Affordance perception aims to localize actionable regions supporting embodied interaction, yet 2D and 3D affordance grounding have evolved as separat...
By Yuhao Liu, Yiming Zhong, Hanqing Wang, Shaocheng Yan, Yuhang Zhang, Wenzhou Lyu, Ziyang Ding, Wei Zhang, Xue Zhao, Jin Pan, Yuexin Ma, Xinge Zhu
arXiv:2609.37810v1 Announce Type: cross
Abstract: Vision-language-action and world-action models have demonstrated impressive capabilities in robotics, yet generalization to unseen tasks remains chal...
By Sicheng Xie, Yitong Chen, Haidong Cao, Shunlin Lu, Zuxuan Wu, Yu-Gang Jiang