arXiv AI By Eunjae Jo, Nakyung Lee, Gyuyeong Kim

Database Normalization via Dual-LLM Self-Refinement

Read the original on arXiv AI →

arXiv:2508. 17693v2 Announce Type: replace-cross Abstract: Database normalization is crucial to preserving data integrity.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

arXiv Computation and Language
Sep 11

Can LLMs Normalize Databases? A Benchmark and Multi-Agent Framework for Schema Normalization

The paper introduces a Database Normalization Benchmark (DNBENCH) with 3,275 samples to evaluate how well Large Language Models (LLMs) can perform database normalization from 1NF to BCNF, assessing semantic equivalence, structural accuracy, and logical validity. It identifies common failures in dependency inference, schema decomposition, and inter-table constraint reconstruction across various complexity levels. The authors also propose a Multi-Agent Reasoning for Schemas (MARS) framework that separates evidence extraction, violation diagnosis, and decomposition planning from schema generation, achieving an 82.0% improvement in DNB-SCORE over a single-prompt baseline.

By Dong-Jae Koh, Huisu Kim, SeongHwan Yoon, Lasse M. Jantsch, Chun-Hee Lee, Seonghyeon Lee, Young-Kyoon Suh
arXiv AI
Aug 18

ACTS-SQL: Agentic and Critic-Oriented Tree-Structured SQL Correctness with Large Language Models

arXiv:2608. 15145v1 Announce Type: new Abstract: Large Language Models (LLMs) have been increasingly adopted in Text-to-SQL systems, yet SQL errors remain a major obstacle in real-world Text-to-SQL inference pipelines.

By Xinmei Huang, Jie Song, Peng Li, Fuxin Jiang, Jing Zhang, Tieying Zhang, Jianjun Chen, Chenming Liu, Tao Yang, Maoyin Liu, Wenda Li, Hong Chen, Cuiping Li
arXiv AI
Sep 7

A Cost-Aware Agentic Architecture for NL-to-SQL over Nested Enterprise Schemas, with a New Benchmark

The paper introduces the DevRev NL2SQL benchmark, featuring 900 execution‑verified queries that test natural‑language‑to‑SQL systems on nested, graph‑like enterprise schemas, and proposes the Semantic Depth Score (SDS) as a rubric for analytical reasoning depth. It also presents a cost‑aware, single‑generation agentic architecture that includes schema selection, metadata retrieval, and error‑repair components tailored to these complex schemas. On the DevRev benchmark, the system achieves 91.7% answer correctness, outperforming the next‑best baseline by 54.6 percentage points, and remains competitive on the Spider 2.0 Snowflake dataset.

By Yoga Sri Varshan Varadharajan, Ajay Yadav, Ritesh Goru, Prateek Chaudhury, Constantine Caramanis, Prateek Jain, Divyateja Pasupuleti, Sunil Kumar Pandey