arXiv AI By Junjie Pang, Zhenzhen Xie, Haoke Han, Ying He, Jing Wang, Gang Liu

DNative-Twin: Decision Graphs and Digital Twins for Reconstructable Agentic Decisions

Read the original on arXiv AI →

DNative‑Twin is a graph‑native digital twin that records an AI agent’s committed decision as a typed trajectory, linking observed state, decision path, and authority. It re‑executes the decision mechanism under declared conditions, synchronizing and replaying the process in isolation to compare outcomes under controlled changes. Experiments on enterprise decision logs show that adding replay‑contract state and verification evidence improves unresolved‑divergence recall from 0 to 1.0, while end‑to‑end processing time rises from 0.794 to 8.889 seconds across 500–5,000 cases.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

arXiv AI
Sep 24

TwinCheck: Evidence-Grounded Negative-Twin Verification for Stateful Tool Agents

TwinCheck is an inference‑time verification policy for stateful tool agents that only replaces a proposed tool call when a trace‑grounded counterfactual alternative, called a negative twin, satisfies structural checks and is preferred by a pairwise verifier in both candidate orders. The method uses exact replay to isolate intervention effects, and in experiments on 159 multi‑turn BFCL V4 tasks, it increased GPT‑5.6 Sol’s task success from 45.3% to 58.5% without any observed success‑to‑failure regressions.

By Jiaxuan Dai, Tianyi Huang