arXiv:2605. 26397v2 Announce Type: replace-cross Abstract: Safety alignment reduces explicitly harmful outputs but inadvertently encodes a sanitized, neuronormative representation of marginalized communication.
By Naba Rizvi, Mohammed Rizvi, Harper Strickland, Saleha Ahmedi, Nedjma Ousidhoum
arXiv:2609.15369v1 Announce Type: new
Abstract: Word-level detectors identify unedited AI-generated text almost perfectly, but the literature documents their brittleness under rewording, and a word-l...
By Jochen Madler (Sitefire)
arXiv:2606. 04906v1 Announce Type: cross Abstract: Although it is generally agreed that AI-generated text poses a broad societal risk, there is no common understanding in the AI-generated text detection literature on what constitutes harmful use.
By Nils Dycke, Marina Sakharova, Nico Daheim, Iryna Gurevych
arXiv:2603.15034v2 Announce Type: replace-cross
Abstract: This paper replicates and extends the system used in the AuTexTification shared task for authorship attribution of machine-generated texts. E...
By Adam Skurla, Dominik Macko, Jakub Simko
arXiv:2607. 26062v1 Announce Type: cross Abstract: Background: This work investigates the presence of implicit bias in Large Language Model (LLM)-based chat AI models directed toward people with intellectual disabilities (ID).
By Karly V. Coffey, Gloria L. Krahn, John P. Hanley, Jacob E. Neely
arXiv:2606. 12073v1 Announce Type: cross Abstract: Generative AI has made fluent prose cheap to produce, breaking the old promise to readers that good writing meant real thinking.
By Jason Miklian, John E. Katsos