arXiv AI By Cagri Temel

CT-SAFR: Safe and Interpretable Chain-of-Thought Reasoning for Autonomous Robots: A Multi-Layered Verification Framework for Trustworthy AI-Driven Robotic Decision Making

Read the original on arXiv AI →

CT‑SAFR is a multi‑layered verification framework designed to enhance the safety and faithfulness of Chain‑of‑Thought reasoning in autonomous robots. The framework achieves a 94.2% hallucination detection rate with sub‑500 ms latency, and a warehouse robot case study shows an 87% reduction in unsafe reasoning outputs. The study also offers recommendations for responsible deployment of reasoning‑capable autonomous robots.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

arXiv AI
Sep 3

Towards Trustworthy Autonomous Robots: An Explainable AI-Based Decision Framework

The paper introduces TRACE, a Transparent Reasoning Architecture for Credible Execution, which provides an explainable AI-based decision framework for autonomous robots. TRACE structures decision-making into four auditable layers—Semantic Perception, Belief Reasoning, Action Synthesis, and Execution Verification—to ensure every action can be traced back to sensor evidence through documented causal chains. Experimental results on warehouse robot navigation show high evidence traceability (98.6%), temporal continuity (99.0%), and decision reconstructability (98.1%) across 500 simulated decision cycles.

By Cagri Temel
arXiv Machine Learning
Jun 24

Verifiable Foundation Models for Robot Safety

arXiv:2606. 23754v1 Announce Type: cross Abstract: Deploying foundation models for robot control raises a central challenge: the expressive power that enables rich, multimodal perception also makes these models opaque and difficult to analyze formally, rendering them intractable for existing verification tools.

By Davide Corsi, Kyungmin Kim, Roy Fox
arXiv AI
4d ago

Bridging Thought and Action: Taming Long-Horizon Instability in Open-Source LLM Agents with a MetaTool-Enhanced ROS Framework

The paper introduces a ROS-Agent architecture that enhances task reliability and execution efficiency for open‑source LLM‑powered robotic agents. It adds a MetaTool that forces the LLM to produce a structured pseudo‑code plan before any action, storing this plan in a scratchpad to separate planning from execution. Experiments on a custom mobile robot show up to ~24% improvement in complex task completion and contextual consistency compared to the baseline.

By Kazi Abrar Mahmud, Nilotpaul Kundu Dhurubo, Tamal Kirttonia, Sabbir Hossain Ujjal, Mohammad Ariful Haque