STAGE: Stateful Translation to Agentic Graph Execution with Policy-Scoped Context and Deterministic Control
Read the original on Hugging Face Trending Papers →The Flow has not summarised this story yet — read it at Hugging Face Trending Papers.
The Flow has not summarised this story yet — read it at Hugging Face Trending Papers.
arXiv:2608.22538v1 Announce Type: new Abstract: Policy-governed agents must interpret case evidence while following an authorized procedure. We present \textsc{Stage}, an executable-graph framework t...
The paper introduces a governed approach to enterprise analytics in which a language model interprets user queries and a deterministic policy selects and runs pre‑approved analytical programs that return both results and evidence. The authors demonstrate that this restriction remains expressive for a defined analytical class—including relational operations, aggregation, comparison, windows, ranking, and similarity—while ensuring reproducibility through fixed meaning, policy, data, and execution rules. In experiments with 440 runs, three 8B models generated SQL and selected tools at runtime, whereas a policy‑executed analyzer achieved a perfect 110/110 match across all test datasets, though no runtime‑planning episodes matched the full answer‑and‑evidence contract. "whyItMatters":"The study shows that a governed, policy‑driven framework can reliably produce accurate, reproducible analytics results, highlighting a viable path for controlled AI‑driven data analysis."
arXiv:2609.14400v1 Announce Type: new Abstract: Agent benchmarks evaluate policy compliance but assume each policy determines a unique correct action. Natural-language policies can violate this assum...
arXiv:2609.37457v1 Announce Type: new Abstract: Enterprise artificial-intelligence agents increasingly call tools, modify infrastructure, and process protected data, creating a need to separate actio...
arXiv:2608.29971v1 Announce Type: new Abstract: As agentic systems evolve into complex multi agent orchestration workflows, there is a growing and critical need for systematic frameworks that measure...
The paper introduces Aegis, a runtime governance system for agentic AI that treats model outputs as action proposals and mediates them through a trusted decision layer before tool execution. Aegis evaluates proposals against active policy, resolves provenance server‑side, fails closed under uncertainty, and routes selected cases through a Senate‑style settlement process. In a sandbox evaluation across 6,300 rows, Aegis prevented all governed mock‑tool applications and risky side‑effect completions, preserving provenance and quorum evidence for all settled cases.