arXiv AI By Zhe Sun, Meng Wang, Lei Wang, Yuxi Wang, Wanxin Li, Yujia Peng, Zhenliang Zhang

BotDirector: Robot Storytelling Across the Symmetrical Reality with Multi-modal Interactions

Read the original on arXiv AI →

arXiv:2606. 03223v1 Announce Type: cross Abstract: Robot storytelling offers a unique blend of technological innovation and creative expression that engages children in unprecedented ways.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

arXiv Machine Learning
Aug 10

Robot guide with multi-agent control and automatic scenario generation with LLM

arXiv:2509. 10317v2 Announce Type: replace-cross Abstract: The article describes the development of a hybrid social robot control architecture to overcome the limitations of traditional approaches, where behavior scripts manually synchronize the robot's actions and text, and existing methods focus primarily on short dialogue responses.

By Elizaveta D. Moskovskaya, Anton D. Moscowsky
arXiv AI
Aug 13

Empowering Children to Create AI-Enabled Augmented Reality Experiences

arXiv:2508. 08467v2 Announce Type: replace-cross Abstract: Despite their potential to enhance children's learning experiences, AI-enabled AR technologies are predominantly used in ways that position children as consumers rather than creators.

By Lei Zhang, Shuyao Zhou, Amna Liaqat, Tinney Mak, Brian Berengard, Emily Qian, Andr\'es Monroy-Hern\'andez
arXiv Computer Vision
Aug 28

4DSynth: Controllable Procedural World Synthesis for Dynamic Embodied Simulation

4DSynth is a controllable procedural system that transforms natural-language descriptions, blueprint masks, or single photographs into editable 4D environments featuring explicit geometry, animated actors, collision-free trajectories, and physics-ready simulation states. The system unifies animation, camera planning, rendering, and task generation within a single geometry-grounded representation, enabling scalable creation of dynamic embodied simulation scenes. Using 4DSynth, the authors built 4DSynth-Nav, an interactive navigation benchmark that demonstrates the reproducibility and tunability of procedural failures across vision‑language models.

By Zehao Qi, Haochen Luo, Jia-Wang Bian, Zeyu Ma, Shuyang Sun