Position: Preventing AI-Generated CSAM Necessitates New Approaches to AI Safety
arXiv:2607. 05407v1 Announce Type: cross Abstract: Modern artificial intelligence (AI) systems present profound new risks to child safety.
Researchers developed an auditing technique to test generative AI models for malicious capabilities, without prompting them for illegal outputs.
arXiv:2607. 05407v1 Announce Type: cross Abstract: Modern artificial intelligence (AI) systems present profound new risks to child safety.
Our latest report featuring case studies of how we’re detecting and preventing malicious uses of AI.
arXiv:2506. 06488v3 Announce Type: replace Abstract: A key tool in developing safe AI models is \emph{data auditing}, i.
arXiv:2502. 12445v2 Announce Type: replace Abstract: AI safety is a rapidly growing area of research that seeks to prevent the harm and misuse of frontier AI technology, particularly with respect to generative AI (GenAI) tools that are capable of creating realistic and high-quality content through text prompts.
arXiv:2607. 13453v1 Announce Type: cross Abstract: Artificial Intelligence (AI), especially Generative AI (GenAI), adoption has increased in industries significantly in recent years.
arXiv:2501. 14728v2 Announce Type: replace-cross Abstract: While generative artificial intelligence (GenAI) models have achieved significant success, their misuse for generating deceptive content raises growing concerns about online information security.
arXiv:2410. 01574v4 Announce Type: replace-cross Abstract: The rapid advancement of Generative Artificial Intelligence (GenAI) capabilities is accompanied by a concerning rise in its misuse.
Our latest threat report examines how malicious actors combine AI models with websites and social platforms—and what it means for detection and defense.
arXiv:2606. 05647v1 Announce Type: new Abstract: AI coding agents are increasingly embedded in real-world software development, collaborating with human developers while gaining broader access to codebases and tools.
arXiv:2608. 14730v1 Announce Type: cross Abstract: The rapid evolution of visual generative AI has introduced a wide range of intellectual property risks, spanning the unauthorized learning, reproduction, extraction, misuse, and redistribution of protected data and model assets.
arXiv:2408. 02379v2 Announce Type: replace-cross Abstract: Developing and certifying safe - or so-called trustworthy - AI has become an increasingly salient issue, especially in light of upcoming regulation such as the EU AI Act.