arXiv Machine Learning By Nicholas Chandler, Sebastian J\"ager, Philipp Jung, Felix Bie{\ss}mann

CURED: Creating, Understanding, and Repairing Errors Demonstrator

Read the original on arXiv Machine Learning →

arXiv:2607. 20140v1 Announce Type: new Abstract: Detecting and cleaning errors in tabular data is a prerequisite for data intense software applications.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv Machine Learning.

arXiv Machine Learning
Aug 4

LakeMLB: Data Lake Machine Learning Benchmark

arXiv:2602. 10441v2 Announce Type: replace Abstract: Data lakes have become a fundamental platform for large-scale machine learning by enabling flexible management of heterogeneous data.

By Feiyu Pan, Tianbin Zhang, Aoqian Zhang, Yu Sun, Zheng Wang, Lixing Chen, Li Pan, Jianhua Li
arXiv AI
Jun 16

Understanding, Detecting, and Repairing Real-World In-Context-Learning-Based Text-to-SQL Errors

arXiv:2501. 09310v3 Announce Type: replace-cross Abstract: Large language models (LLMs) have been adopted for text-to-SQL tasks, utilizing their in-context learning (ICL) capability to translate natural language questions into SQL queries.

By Jiawei Shen, Chengcheng Wan, Ruoyi Qiao, Jiazhen Zou, Hang Xu, Yuchen Shao, Yueling Zhang, Weikai Miao, Geguang Pu