arXiv AI By Rishanth Rajendhran, Jenna Russell, Mohit Iyyer, John Frederick Wieting

POLARIS: Guiding Small Models to Write Long Stories

Read the original on arXiv AI →

arXiv:2606. 04095v1 Announce Type: cross Abstract: Small open-weight models struggle at long-form creative writing: their generated stories either fall far short of the requested length, or their quality significantly degrades as length increases, especially when compared to frontier models.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

arXiv Computation and Language
Sep 7

Generating Constructive Feedback on Stories via Reinforcement Learning

The paper introduces a reinforcement learning method to improve the quality of feedback generated by large language models for creative writers. By training with group relative policy optimization and a multi‑component reward that emphasizes tailored, actionable, and critical‑issue‑focused feedback, the authors demonstrate that their approach outperforms existing LLMs and baselines on three story corpora. The study shows that actionable suggestions are the key factor driving constructive feedback.

By Maja Stahl, Timon Ziegenbein, Henning Wachsmuth
arXiv AI
Aug 28

A Multi-Framework Comparison of Outline Stages in Long-Form Generation with LLMs

The paper presents a benchmark that compares seven long‑form generation frameworks across three granularities—single chapter, multi‑chapter, and whole book—using an anchor‑based LLM‑as‑a‑judge protocol to evaluate outlines directly. Results show no single framework dominates across all settings; performance depends on how well a framework’s output form matches the target granularity, with SuperWriter excelling in length‑constrained single‑chapter mode but losing advantage in whole‑book mode. The study finds only moderate correlation between outline and writing quality, supporting the idea that these two stages should be evaluated separately.

By Yifan Song