arXiv:2606. 28749v1 Announce Type: cross Abstract: Although most undergraduates now use large language models (LLMs), a form of generative artificial intelligence (GenAI) for academic writing, no validated method distinguishes the qualitatively different ways students rely on them.
By Shahin Hossain
Large language models (LLMs) are widely used to assist writing, but this study shows they alter both tone and meaning of human text. A user study found that heavy LLM use increased neutral essays by nearly 70% and made writers feel less creative and less in their voice. Even when prompted to make only grammar edits, LLMs changed the semantic content of essays and produced AI-generated scientific reviews that were less focused on clarity and significance and scored higher on average.
By Marwa Abdulhai, Isadora White, Yanming Wan, Ibrahim Qureshi, Joel Z. Leibo, Max Kleiman-Weiner, Natasha Jaques
arXiv:2608.28837v1 Announce Type: cross
Abstract: We conducted an interview study with twelve students on their use of generative AI in academic communication. Students delegated professional message...
By Jared Ren, Soobin Cho
The study investigates whether AI assistance leaves a temporal fingerprint in writing and programming tasks. By analyzing keystroke-level data from three corpora, the authors find that AI contributions appear in distinct bursts and that temporal patterns can almost perfectly distinguish wholesale delegation from authentic work, though ordinary collaboration remains hard to detect. The research suggests that process visibility could serve as a basis for academic integrity checks.
By Eduardo Davalos, Yike Zhang
arXiv:2609.36544v1 Announce Type: cross
Abstract: Generative AI has changed how students produce writing assignments. The final artifact is no longer sufficient to understand the process through whic...
By Divyansh Chandarana, Sandipan De, Vivek Gupta
arXiv:2310.00436v2 Announce Type: replace
Abstract: Authorship identification uses patterns in writing to infer who wrote a text, but those patterns also reflect topic, genre, and register. This surv...
By Haining Wang
arXiv:2603. 04982v3 Announce Type: replace-cross Abstract: Can targeted user training unlock the productive potential of generative artificial intelligence in professional settings?
By Benjamin M. Chen, Hong Bao
The paper introduces SWIM, a task that frames student writing simulation as proficiency‑conditioned essay generation. It evaluates prompting, supervised fine‑tuning, and reinforcement learning for aligning generated essays with student proficiency profiles, using automated essay scoring as a metric. Results show that prompting alone offers limited control, while supervised fine‑tuning and reinforcement learning significantly improve alignment across content, lexical, grammatical, and organizational traits, though low‑proficiency writing remains difficult to replicate.
By Heejin Do, Jakub Kontak, Mrinmaya Sachan
The study explores how undergraduate computing students in Saudi Arabia perceive AI‑generated writing feedback when they are explicitly told that ChatGPT, not a human instructor, produced the score and comments. Through qualitative reflections, four themes emerged: students found the feedback useful for surface‑level revisions, recognized AI’s contextual and pedagogical limits, trusted the feedback conditionally—separating its utility from its authority—and reaffirmed the human instructor’s role as the ultimate grading authority. The findings highlight a clear distinction students make between feedback usefulness and evaluative authority, treating them as separate judgments rather than opposing ends of a single approval scale.
By Rayed AlGhamdi
arXiv:2510.08831v2 Announce Type: replace
Abstract: As AI writing tools become widespread, we need to understand how both humans and machines evaluate literary style, a domain where objective standar...
By Wouter Haverals, Meredith Martin
The study investigates how generative AI (GenAI) affects student learning in AI-related courses, using survey data from 118 students across 12 courses. Four distinct user clusters were identified—high-use, light-use, and two moderate-use groups—each showing varying benefits and reliance patterns. The research highlights that early reliance, evaluation literacy, and instructor policies significantly influence perceived academic benefits and negative impacts, underscoring the need for institutional policies to address inequities in AI use.
By Lydia Manikonda, Mei Si, Sirajam Munira, Oshani Seneviratne, Kristin Bennett
The paper examines whether state‑of‑the‑art large language models (LLMs) produce feedback that aligns with expert teachers’ pedagogical practices, focusing on feedback type and adaptivity. Using a refined taxonomy of seven feedback focus types, the authors annotate and compare teacher and LLM‑generated feedback from three university writing courses, creating the FeedType benchmark. Their analysis shows that while most LLMs cover many feedback types, they do not match teachers’ distribution patterns or adaptive behavior across draft stages and student performance levels.
By Norah Almousa, Shayan Peyghambari Oskoui, Raquel Coelho, Gayle Rogers, Xiang Lorraine Li, Diane Litman