arXiv AI By Jean-Baptiste Espinasse (DiverSe), Djamel Eddine Khelladi (DiverSe, CNRS, IRISA, KHORA), Mathieu Acher (INSA Rennes, IRISA, DiverSe)

Spec2COBOLRot: An Agentic-AI Degradation Loop for Realistic COBOL Corpus Generation

Read the original on arXiv AI →

The paper introduces Spec2COBOLRot, an agentic AI pipeline that generates realistic COBOL programs by combining specification-driven creation with iterative degradation guided by real production code patterns and complexity targets. The authors evaluate the pipeline on three programs from different business domains, showing that it reliably produces syntactically valid code and increases structural complexity, but it does not consistently preserve business behavior. They discuss the limitations of targeting structural metrics alone and propose future work that would generate legacy programs from scratch along a simulated development history.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

arXiv AI
Jun 8

EvoClaw: Evaluating AI Agents on Continuous Software Evolution

arXiv:2603. 13428v2 Announce Type: replace-cross Abstract: With AI agents increasingly deployed as long-running systems, it becomes essential to autonomously construct and continuously evolve customized software to enable interaction within dynamic environments.

By Gangda Deng, Zhaoling Chen, Zhongming Yu, Haoyang Fan, Yuhong Liu, Yuxin Yang, Dhruv Parikh, Rajgopal Kannan, Le Cong, Mengdi Wang, Qian Zhang, Viktor Prasanna, Xiangru Tang, Xingyao Wang
arXiv AI
Aug 24

SDAD: Spec-Driven Agentic Development for the AI-Native SDLC

The paper introduces Spec-Driven Agentic Development (SDAD), a framework that leverages large language models to ingest extensive functional requirement documents and repository context in a single workflow, turning specification quality into the engine for autonomous software delivery. SDAD blends disciplined upfront formalisation with rapid implementation, encompassing intent capture, machine‑readable specifications, agentic synthesis, and multi‑agent verification with human sign‑off. It positions AI‑code as a fourth production paradigm, compares it to traditional Waterfall and Agile approaches, and extends the model to team role evolution, quantitative governance metrics, and a staged migration blueprint for practical adoption.

By Vu Hung Nguyen, Thanh Nguyen