Octopus Protocol is a hardware onboarding framework that allows an AI coding agent to automatically discover, identify, and integrate hardware devices into an AI system. Using a single bootstrap command, the agent runs a five-stage pipeline to enumerate visible hardware, infer device capabilities, generate typed Model Context Protocol tools, produce the necessary code, and activate a live endpoint. The system maintains a persistent daemon that repairs deployment failures, enabling consistent, platform‑agnostic interfaces across diverse hosts without manual integration code.
By Quilee Simeon, Justin M. Wei, Yile Fan
CUA-Universe is a scalable environment-to-data pipeline that transforms real desktop software into hybrid GUI+CLI environments, enabling agents to coordinate visual inspection with command-line operations. It includes App-Forge for reproducible VMs and CLI surfaces, Task-Weave for generating diverse hybrid tasks, and Path-Steer for efficient rollouts and trajectory harvesting. Training on this data improves agent success and efficiency across multiple benchmarks, demonstrating the value of hybrid interaction.
By Haoting Shi, Wenhao Wang, Weicheng Fang, Yaozhong Liang, Tian Jin, Pengxiang Zhao, Guangyi Liu, Siheng Chen, Yanfeng Wang
arXiv:2608. 16178v1 Announce Type: cross Abstract: Operational telemetry is predominantly engineered for human reading: systems repeatedly serialize verbose prose, static keys, and redundant context across billions of log lines.
By Jun He, Deying Yu
arXiv:2509. 17255v2 Announce Type: replace-cross Abstract: We present the first language-model-driven agentic artificial intelligence (AI) system to autonomously execute multi-stage physics experiments on a production synchrotron light source.
By Thorsten Hellert, Drew Bertwistle, Simon C. Leemann, Antonin Sulc, Marco Venturini
LabFactory is a framework that transforms a scientific brief into an executable AI lab, integrating models, knowledge resources, tools, and a controller behind a fixed interface. The builder packages the lab in a metered workspace, and a separate host evaluates the delivered artifact on held‑out inputs, ensuring the system itself is the evaluation target. Across 28 constructions in seven scientific domains, the delivered labs surpassed reference values on all 33 subtests, demonstrating that an AI agent can fully realize a scientific brief into a working, inspectable lab.
By Jinge Wu, Hongjian Zhou, Mingde Zeng, Jiayuan Zhu, Junde Wu, Jiazhen Pan, Lei Clifton, Andrew Liu, David A. Clifton
arXiv:2606. 03755v1 Announce Type: new Abstract: Autonomous science is moving from demonstration to infrastructure.
By Linwu Zhu, Liqiang Gao, Yan Chen, Dan Zhu, Jian Huang