The paper introduces ToW3D, a method for precise and consistent control over 3D generative adversarial networks (GANs) using a Tug-of-War approach between shape deformation and appearance consistency. It addresses the challenge that 3D generators often lack generalization, leading to drastic global appearance changes when editing local mesh areas. ToW3D employs a two-step optimization—drag locally and shove globally—along with a structure adaptation module and a semantic preservation module, achieving superior appearance consistency and fidelity compared to prior methods, especially under large deformations.
By Haixu Song, Fangfu Liu, Chenyu Zhang, Yueqi Duan
The paper introduces NPLSD, a pair of line‑segment detectors optimized for the Neural‑ART NPU on STM32N6 microcontrollers. NPLSD‑H retains a convolutional backbone and replaces transformer components with a fully‑convolutional head, while NPLSD‑M adapts a lightweight trunk to the NPU’s supported operators. Trained on ImageNet and ShanghaiTech Wireframe, the models achieve competitive accuracy with 2.63 M and 0.62 M parameters, respectively, and ablation studies show initialization contributes significantly to performance.
By Parsa Hassani Shariat Panahi, Amir Hossein Jalilvand, M. Hassan Najafi
arXiv:2609.25511v1 Announce Type: cross
Abstract: Vision-based gesture control accepts commands from any hand in the camera field of view, which is unsafe in shared indoor spaces. This paper presents...
By Diyari Mohammed Salih, Ilyes Chaabeni, Naima Ait Oufroukh
arXiv:2609.26420v1 Announce Type: cross
Abstract: Text-to-motion models generate plausible human motion but do not model a robot's dynamics; whole-body tracking controllers execute robot references r...
By Raphael Memmesheimer, Sven Behnke
arXiv:2609.26792v1 Announce Type: cross
Abstract: Faithfully evaluating end-to-end driving policies in simulation requires observations that are not merely photo-realistic, but preserve the scene fea...
By Ziyang Leng, Sicheng Mo, Seth Z. Zhao, Haoyuan Cai, Yu Zeng, Rowan McAllister, Bolei Zhou
arXiv:2609.26795v1 Announce Type: cross
Abstract: 3D Gaussian Splatting (3DGS) can reconstruct a captured scene photorealistically, but the resulting representation does not by itself support physica...
By Runyi Yang, Deheng Zhang, Xiaoye Wang, Kanzhi Wu, Lei Sun, Ajad Chhatkuli, Kunyu Peng, Luc Van Gool, Danda Pani Paudel
arXiv:2509.04600v2 Announce Type: replace
Abstract: Reconstructing global human motion from monocular video is fundamental to VR, graphics, and robotics, yet remains ill-posed due to depth ambiguity,...
By Zhongyuan Hu, Qijun Ying, Jiazhi Shu, Ronghui Li, Yu Lu, Zijiao Zeng, Xiu Li
arXiv:2609.22691v1 Announce Type: new
Abstract: An enduring and richly elaborated dichotomy in cognitive neuroscience is that of human behavior control mechanisms, divided into habitual versus goal-d...
By Chongyu Bao, Haokai Yang, Yuhan Wang, Zhaochong An, Kunpeng Liu, Xiaolan Liu
arXiv:2609.23695v1 Announce Type: new
Abstract: Recent advances in Physical AI have accelerated the use of foundation models in autonomous systems such as unmanned aerial vehicles (UAVs), which must...
By Mohamed Amine Ferrag, Merouane Debbah, Abderrahmane Lakas, Manu Perumkunnil, Norbert Tihanyi
arXiv:2609.23974v1 Announce Type: new
Abstract: Foundation models are endowing autonomous systems with greater intelligence, enabling a more comprehensive understanding of the environment through vis...
By Boxun Hu, Jiawei Ge, Axel Krieger, Peng Wang, Tinoosh Mohsenin
arXiv:2609.22538v1 Announce Type: cross
Abstract: Large language model (LLM) planners can decompose natural-language instructions and select reusable robot skills, but choosing the correct skill does...
By Ajay Vikram Periasami, Xinyuan Luo, Haoyu Li, Xianyi Cheng
arXiv:2609.22609v1 Announce Type: cross
Abstract: Insertion is a fundamental operation in robotic construction assembly, where variations in material properties and assembly conditions make it diffic...
By Lin He, Yanyi Chen, Haofei Sun, Lingyao Li, Min Deng
arXiv:2609.22611v1 Announce Type: cross
Abstract: Humanoid robots can acquire complex skills by imitating kinematic humanoid motion references, yet reliable references for contact-rich interactions r...
By Lalit Jayanti, Kashu Yamazaki, Yuto Shibata, Kotaro Amaya, Katerina Fragkiadaki
arXiv:2609.23407v1 Announce Type: cross
Abstract: Humans can effortlessly localize the direction of a sound source and integrate it with visual cues for reasoning, yet this remains challenging for em...
By Ruixun Liu, Yuxuan Wang, Jiacheng Xie, Yuhuan You, Donghua Cai, Junming Lin, Xiong-Hui Chen, Zhifang Guo, Yunfei Chu, Qize Yang, Xize Cheng, Jin Xu, Yiwu Zhong
arXiv:2609.24274v1 Announce Type: cross
Abstract: Deploying language-conditioned manipulation without a dedicated GPU requires efficient inference and action chunks that cover the delay between polic...
By Khanh D. Nguyen, Hoang M. Truong, An T. Le
arXiv:2609.24369v1 Announce Type: cross
Abstract: Deception plays a central role in Intelligence operations, yet it remains difficult to analyse systematically without expert knowledge of reasoning p...
By Stefan Sarkadi, Xabier Garmendia, Jack Mumford, Trevor Bench-Capon
arXiv:2510.26915v2 Announce Type: replace-cross
Abstract: While heterogeneous teams have typically been designed for well-specified missions with known semantics, generative intelligence, i.e., large...
By Zachary Ravichandran, Fernando Cladera, Ankit Prabhu, Jason Hughes, Carlos Nieto-Granda, Varun Murali, Camillo Taylor, George J. Pappas, Vijay Kumar
arXiv:2603.03768v2 Announce Type: replace-cross
Abstract: Full-stack human-robot collaboration (HRC) can become brittle when replacing a planner, partner model, coordination policy, or controller cha...
By Hao Zhang, Yisen Li, Ruize Geng, Yves Tseng, Yaru Niu, Ding Zhao, H. Eric Tseng
arXiv:2606.23079v2 Announce Type: replace-cross
Abstract: Neural world models coupled with model predictive control (MPC) replan at every environment step to bound accumulated prediction error, but t...
By Yutian Cheng, Xiaojian Ma, Xianhao Wang, Min Yang, Rongpeng Su, Hangxin Liu, Xi Chen, Shuai Li, Qing Li
arXiv:2608.02270v2 Announce Type: replace-cross
Abstract: Agriculture 4.0 robotic systems improve field efficiency yet remain too capital-intensive for the fragmented smallholdings that dominate glob...
By Weijie Shi, Zicheng Xu, Zhenbang Cheng, Haoran Xuan, Mingbo Duan, Gan Ge