arXiv AI By Florian St\"ortz, Catalin-Andrei Stan, Alexandru Dinu, Sandra Servia-Rodr\'iguez, Mihaela Gaman, Calin Miron, Edward Raff

Large Byte Model: Teaching Language Models About Compiled Code

Read the original on arXiv AI →

arXiv:2606. 02834v1 Announce Type: cross Abstract: Malware analysis starts with the raw bytes of an executable program, and tools to "lift" these to higher-level representations, such as assembly, are expensive and subject to error.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

arXiv Machine Learning
Sep 18

ALIBI: Adversarial Legitimacy Injection in Binary Input against LLM Malware Analyzers

The paper introduces ALIBI, a semantic cover story attack that injects a small, non-executed read‑only section into compiled binaries to mislead large language model (LLM) malware analyzers. By embedding a coherent but false security narrative, ALIBI can cause LLMs such as Gemini 2.5 Pro, GPT‑5.5 Pro, and Claude Opus 4.7 to downgrade or flip the verdicts of malicious samples. The attack also transfers to ELF binaries, and even a verification‑guided defense prompt only partially mitigates the effect, leaving a significant portion of malicious samples classified as benign.

By Hyeongjun Choi, Wonyoung Jung, Haehoon Seo, Sungyup Nam