arXiv:2509. 14474v3 Announce Type: replace Abstract: The debate around Artificial General Intelligence (AGI) remains open due to two fundamentally different goals: replicating human-level performance versus replicating human-like cognitive processes.
By Meltem Subasioglu, Nevzat Subasioglu
Benchy is a semantic language and execution engine designed to standardize task-oriented AI benchmarks. Each benchmark is fully defined by a program, a scoring function, and a dataset (B=(P,S,D)), and is independent of the AI system that runs it. Benchmarks are authored in canonical YAML, compiled deterministically into JSON, and executed via a universal runtime contract that exposes a named-field input object and a named-field output object, ensuring consistent integration across AI systems.
By Francis F Daniel, Mauro Iba\~nez, Francis Perelman, Marian Basti
arXiv:2606. 10106v1 Announce Type: cross Abstract: The term agent harness now circulates widely in software engineering with generative artificial intelligence.
By Sanderson Oliveira de Macedo
arXiv:2608. 20201v1 Announce Type: new Abstract: Software form has undergone two paradigm shifts since its inception: Software 1.
By Wei Lin, Tao Zhou, Zhaofei Xie, Changgui Hong
arXiv:2606. 05608v2 Announce Type: replace-cross Abstract: For over half a century, software engineering has operated on a foundational premise: human engineers decompose problems, encode decision logic into static code, and manually adapt that code as requirements evolve.
By Zhenfeng Cao
arXiv:2605. 22093v3 Announce Type: replace Abstract: Knowledge graphs have become the primary vehicle for data integration and are critical to the success of modern AI, but the diversity of KG modelling practices, from lightweight vocabularies to richly axiomatised ontologies, makes integration and reuse expensive and brittle.
By Enrico Daga, Valentina Tamma, Terry Payne