arXiv AI

Lost in Transmission: An Information-Theoretic Account of Unsupervised Software Traceability

arXiv:2412. 04704v2 Announce Type: replace-cross Abstract: Traceability remains a critical capability to ensure system reliability, maintainability, and compliance in modern software development.

arXiv AI
Jun 17

Trust-Aware Multi-Agent Traceability: Confidence-Calibrated Knowledge Graphs for Consistent Software Artifact Management

arXiv:2606. 17203v1 Announce Type: cross Abstract: Multi-agent AI systems are increasingly used to automate software engineering tasks including requirements analysis, architecture design, test generation, and traceability linking.

By Mohamed Essam, Kareem Wael, Azza Hassan, Ahmed Haitham, Mahmoud Soliman, Samer Saber, Ibrahim Habib
arXiv AI
Jul 3

ContextNest: Verifiable Context Governance for Autonomous AI Agent

arXiv:2607. 02116v1 Announce Type: new Abstract: Autonomous AI agents increasingly depend on external knowledge stores, yet most retrieval pipelines provide relevance without durable guarantees of provenance, version identity, integrity, traceability, or point-in-time reconstruction.

By Misha Sulpovar (PromptOwl, LLC), Benn R. Konsynski (Goizueta Business School, Emory University), Qaish Kanchwala (IBM Research), Gabe Goodhart (IBM Research)
arXiv AI
Jul 31

TraceCoder: Explainable and Auditable Code Generation with Position-Key Snippet Versioning

arXiv:2607. 26307v1 Announce Type: new Abstract: Contemporary LLM-based coding agents produce code as black-box outputs: the rationale behind each line is hidden, the evolution of the code through benchmark-driven repair is ephemeral, and post-hoc auditing is impossible.

By Rwaida Alssadi, Muntaser Syed, Balaji Kasula, Lamine Deen, Majed Alotaibi, Mohammed Alghamdi, Tyler Ton, Ali Alqarni, Marius Silaghi
arXiv AI
Aug 20

SIDScope: A Diagnostic Resource for Semantic-ID Interfaces in Generative Recommendation

SIDScope is a diagnostic tool that evaluates Semantic‑ID interfaces used in generative recommendation systems. It normalizes item‑to‑code artifacts, verifies provenance, profiles mapping structure, and compares revisions while tracking path‑to‑item outcomes in generated traces. Using nine tokenizer exports from Amazon and Yelp data, SIDScope shows that interface health depends on multiple signals and reveals gaps in prefix alignment, trace accounting, and mapping refresh effects.

By Jiandong Ding, Huijie Qin, Tiandeng Wu, Yi Cao
Hugging Face Trending Papers
Aug 19

SIDScope: A Diagnostic Resource for Semantic-ID Interfaces in Generative Recommendation

SIDScope is a diagnostic tool that evaluates Semantic-ID mappings used between item tokenizers and generative recommenders. It normalizes artifacts, verifies provenance, profiles mapping structure, compares revisions, and tracks path-to-item outcomes in generated traces. Using data from Amazon and Yelp, SIDScope shows that interface health depends on multiple signals, revealing gaps in prefix alignment, trace accounting, and refresh handling that affect model reuse.

arXiv Computation and Language
3d ago

TRACE: Target-Aware Retrieval, Attributed Evidence, and Contract-Constrained Extraction for LitTraceQA

TRACE is a system designed to bridge the grounding contract gap in LitTraceQA by combining target-aware retrieval, independent typed evidence localization, multimodal table extraction, and schema-driven table construction. It indexes 27,487 papers using multiple representations while preserving question targets, predicts observation units for tables, and assembles rows with evaluator-compatible key normalization. On the official test set, TRACE achieves a 0.760613 overall score, with high paper F1, evidence F1, and multiple-choice accuracy, though table-row and macro cell performance remain lower.

By Sachin Gupta, Divya Godara