arXiv AI

CT-SAFR: Safe and Interpretable Chain-of-Thought Reasoning for Autonomous Robots: A Multi-Layered Verification Framework for Trustworthy AI-Driven Robotic Decision Making

CT‑SAFR is a multi‑layered verification framework designed to enhance the safety and faithfulness of Chain‑of‑Thought reasoning in autonomous robots. The framework achieves a 94.2% hallucination detection rate with sub‑500 ms latency, and a warehouse robot case study shows an 87% reduction in unsafe reasoning outputs. The study also offers recommendations for responsible deployment of reasoning‑capable autonomous robots.

arXiv AI
Sep 3

Towards Trustworthy Autonomous Robots: An Explainable AI-Based Decision Framework

The paper introduces TRACE, a Transparent Reasoning Architecture for Credible Execution, which provides an explainable AI-based decision framework for autonomous robots. TRACE structures decision-making into four auditable layers—Semantic Perception, Belief Reasoning, Action Synthesis, and Execution Verification—to ensure every action can be traced back to sensor evidence through documented causal chains. Experimental results on warehouse robot navigation show high evidence traceability (98.6%), temporal continuity (99.0%), and decision reconstructability (98.1%) across 500 simulated decision cycles.

By Cagri Temel
arXiv Machine Learning
Jun 24

Verifiable Foundation Models for Robot Safety

arXiv:2606. 23754v1 Announce Type: cross Abstract: Deploying foundation models for robot control raises a central challenge: the expressive power that enables rich, multimodal perception also makes these models opaque and difficult to analyze formally, rendering them intractable for existing verification tools.

By Davide Corsi, Kyungmin Kim, Roy Fox
arXiv AI
4d ago

Bridging Thought and Action: Taming Long-Horizon Instability in Open-Source LLM Agents with a MetaTool-Enhanced ROS Framework

The paper introduces a ROS-Agent architecture that enhances task reliability and execution efficiency for open‑source LLM‑powered robotic agents. It adds a MetaTool that forces the LLM to produce a structured pseudo‑code plan before any action, storing this plan in a scratchpad to separate planning from execution. Experiments on a custom mobile robot show up to ~24% improvement in complex task completion and contextual consistency compared to the baseline.

By Kazi Abrar Mahmud, Nilotpaul Kundu Dhurubo, Tamal Kirttonia, Sabbir Hossain Ujjal, Mohammad Ariful Haque
arXiv AI
Aug 25

Physical Agentic AI: An Architecture for Orchestrating a Robot Crew with LLMs

Physical Agentic AI proposes an architecture that links semantic planning with physical execution for robot crews. Each robot exposes a typed skill library, while a foundation model planner decomposes tasks into phases and assigns robot‑skill pairs. A Robot Orchestrator validates and authorizes one skill at a time, ensuring actions are grounded in robot capabilities, system state, and workflow constraints before actuation.

By Xinyuan Liu, Eren Sadikoglu, Riana Chatterjee, Ransalu Senanayake
arXiv AI
Jul 14

Think When It Matters: Conditional VLM Reasoning for Social Navigation with RL Policies

arXiv:2607. 10991v1 Announce Type: cross Abstract: As mobile robots become more integrated into everyday human environments, social robot navigation is becoming essential for ensuring human comfort, safety, and trust.

By Ali Ahmadi, Hamed Rahimi, Adrien Jacquet Cretides, Marie Samson, Mahdi Khoramshahi, Mohamed Chetouani