Does AI Help Cyber Attackers or Defenders? Evidence from Nonpublic Vulnerabilities and Subsequent Attacks
Read the original on Hugging Face Trending Papers →The paper investigates how frontier AI systems perform on tasks related to cyber security, specifically exploit generation, vulnerability repair, and subsequent attacks, using five nonpublic software environments. It evaluates both open‑weight and proprietary models, employing deterministic graders rather than LLM judges to score performance. Results show significant variation across systems and vulnerability types, with repair scores higher than attack scores in two environments and lower in three, and highlight that passing an initial security test does not guarantee long‑term defense, as 92 of 524 defender test intervals still experienced successful exploits after the first one was stopped.
Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at Hugging Face Trending Papers.