arXiv Machine Learning By Mian Huang, Xueqin Wang

Elimination Geometry

Read the original on arXiv Machine Learning →

The monograph introduces Elimination Geometry (EG), a typed, native‑loss, audit‑oriented framework that investigates when locally optimal objects can be realized by a shared deployment rule. EG examines how elimination and compression can erase distinctions needed for prediction, inference, control, or representation, and it separates local solvability, global realizability, and finite‑sample certifiability. The work synthesizes tools from geometry, optimization, information theory, statistics, and machine learning to address regular, coordination, singular, compositional, and resource‑limited mechanisms, and demonstrates applications in sparse model selection, distribution‑free prediction, observational treatment policies, routed expert and retrieval systems, and learned score fields.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Machine Learning.

arXiv AI
Aug 5

Don't Regenerate, Debug: A Domain-Specific Agent for Repairing Near-Miss Hardware Operators

arXiv:2608. 02712v1 Announce Type: cross Abstract: Kernel generation for hardware accelerators such as GPUs and NPUs has become a proving ground for large language models (LLMs), and state-of-the-art systems raise correctness through pipelines that couple LLMs with agentic reinforcement learning and evolutionary search.

By Yansong Sun, Shenxiu Wu, Siyuan Chen, Runlin Hou, Junhao Qiu, Junming Cao, Shudi Shao, Zhichao Lu, Qingfu Zhang
arXiv AI
4d ago

From Dead Code and Static Requirements to Working Engines: Software Revival with Coding Agents

The paper introduces ReviveBench, a benchmark designed to evaluate coding agents’ ability to revive non‑running software and reconstruct industrial engines from open specifications. It comprises two families of tasks—revival (ten tasks addressing dependency issues, missing modules, legacy builds, and GPU models) and reconstruction (thirteen tasks covering numerical, geometric, hardware, and transactional systems). The benchmark uses hidden verifiers calibrated against native environments, engineering tools, or reference implementations, and the authors report that the strongest evaluated model passes all revival tasks and most reconstruction tasks, while also uncovering verifier defects that highlight measurement error in executable verification.

By Tianyu Liu, Dingyuan Dai, Yufan Du, Zhen Yang