Skip to content
Review

AutoFiction : Measuring AI ability to execute long-horizon writing tasks

· 0 citations · 45 references

TL;DR

Initial human feedback reveals that AI-written novels contain interesting descriptions and concepts, but often fail in long-range coherence and prose quality, including conceptual repetition, distracting details, and weak dialogues.

View source

Similar papers

Preprint Aug 2026

CraftAlign: Feature-Grounded Evaluation and Revision Guidance for AI Stories

CraftAlign is introduced, a framework that aligns AI stories with the craft of human storytelling by both assessing Human/AI writing patterns and providing revision guidance by both assessing Human/AI writing patterns and providing revision guidance.

Yang Yang, Boyun Xu, Shaofeng Liang et al. · 0 citations
Preprint Aug 2026

SAGE: Self-Evolving Storyboard Skills via Attribution-Guided Rule Evolution

This work presents SAGE (Skill with Attribution-Guided Evolution), a deployed framework that learns, attributes, evolves, and routes directing knowledge from expert demonstrations, and releases PROSE, the first public dataset pairing screenplays with storyboards by professional directors across 68 episodes.

Maolin Ran, Xiaoyan Lu, Jiaqi Liu et al. · 0 citations
Book Open access Jul 2026

A Comparison of Speech and Typing Input for Creative Generative AI Tasks

A user study comparing two modalities for writing prompts for generative AI tasks reveals that input modality significantly influenced prompting behaviour but did not lead to measurable differences in subjective evaluations.

Nishant Rathore, Tushar Billakanti, J. Ceha et al. · 0 citations
Preprint Aug 2026

NARU: A Benchmark for NARrative Evolution and Cultural Nuance Understanding in Japanese Extreme Long Video

NARU, a benchmark designed to evaluate Narrative evolution and Reasoning on cultural Understanding in Japanese long-form video, is introduced, a hierarchical memory-based annotation pipeline that transforms raw video into structured event, narrative, and cultural annotations, then generates questions via task-oriented synthesis and iterative shortcut removal.

Yuheng Huang, Jianlang Chen, Jiayang Song et al. · 0 citations
Preprint Aug 2026

ClueWeaver: Reward-Guided Dual-Agent Evidence Reasoning for Compact LLMs on Literary Long Narratives

Experiments across multiple long-context narrative question answering and claim verification settings show that ClueWeaver substantially improves local end-to-end language models while providing evidence coverage and paragraph-referenced reasoning traces.

Ji-Hao Zhu, Zhiwei Yang, Wenxiao Zhang et al. · 0 citations