arXiv AI

Collective Counterfactual Planning: Coordination, Consent, and Verification under Representational Constraints

arXiv:2608. 17932v1 Announce Type: cross Abstract: Groups routinely complete projects that no single member can plan, execute, or verify alone.

arXiv Machine Learning
Jun 15

Contract-Based Compositional Shielding for Safe Multi-Agent Reinforcement Learning

arXiv:2606. 14130v1 Announce Type: new Abstract: Safe coordination problems surface in multi-agent reinforcement learning when global safety cannot be enforced by any agent unilaterally: the admissibility of one agent's action may depend on the dynamics of other agents.

By Omar Adalat, Edwin Hamel-De le Court, Francesco Belardinelli
Hugging Face Trending Papers
Jun 19

Training the Orchestrator: A Supervised Approach to End-to-End PDDL Planning with LLM Agents

Translating natural-language planning intent into verified plans is a longstanding challenge: people communicate goals in language, while classical planners require formal PDDL specifications. Recent agentic frameworks bridge this gap by orchestrating a pool of specialized repair agents inside a verifier-checked refinement loop, but the orchestrator at the centre is itself a prompted frontier LLM, paying a frontier-LLM API call at every refinement step.