arXiv AI By Zepeng Li, Jie Ren, Zhanyong Tang, Jie Zheng, Zheng Wang

AutoPass: Evidence-Guided LLM Agents for Compiler Performance Tuning

Read the original on arXiv AI →

arXiv:2606. 20373v1 Announce Type: cross Abstract: Large Language Models (LLMs) show promise for code compilation tasks, but applying them to runtime performance tuning is difficult due to complex microarchitectural effects and noisy runtime measurements.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv AI.

arXiv AI
Jun 9

AgentCompile: An LLM-Guided Compiler for Direct CUDA Inference

arXiv:2606. 07665v1 Announce Type: cross Abstract: Transformer inference increasingly depends on specialized compiler and runtime support, but real model graphs still require semantic decisions about which regions are worth specializing and which CUDA implementation families are plausible.

By Xuanzhe Li, Ziyan Weng, Zhiyu Zhu, Junhui Hou