[2604.03136] StoryScope: Investigating idiosyncrasies in AI fiction
Abstract:As AI-generated fiction becomes increasingly prevalent, questions of authorship and originality are becoming central to how written work is evaluated. While most existing work in this space focuses on identifying surface-level signatures of AI writing, we ask instead whether AI-generated stories can be distinguished from human ones without relying on stylistic signals, focusing on discourse-level narrative choices such as character agency and chronological discontinuity. We propose StoryScope, a pipeline that automatically induces a fine-grained, interpretable feature space of discourse-level narrative features across 10 dimensions. We apply StoryScope to a parallel corpus of 10,272 writing prompts, each written by a human author and five LLMs, yielding 61,608 stories, each ~5,000 words, and 304 extracted features per story. Narrative features alone achieve 93.2% macro-F1 for human vs. AI detection and 68.4% macro-F1 for six-way authorship attribution, retaining over 97% of the performance of models that include stylistic cues. A compact set of 30 core narrative features captures much of this signal: AI stories over-explain themes and favor tidy, single-track plots while human stories frame protagonist' choices as more morally ambiguous and have increased temporal complexity. Per-model fingerprint features enable six-way attribution: for example, Claude produces notably flat event escalation, GPT over-indexes on dream sequences, and Gemini defaults to external character description. We find that AI-generated stories cluster in a shared region of narrative space, while human-authored stories exhibit greater diversity. More broadly, these results suggest that differences in underlying narrative construction, not just writing style, can be used to separate human-written original works from AI-generated fiction.
View PDF HTML (experimental) Abstract:As AI-generated fiction becomes increasingly prevalent, questions of authorship and originality are becoming central to how written work is evaluated. While most existing work in this space focuses on identifying surface-level signatures of AI writing, we ask instead whether AI-generated stories can be distinguished from human ones without relying on stylistic signals, focusing on discourse-level narrative choices such as character agency and chronological discontinuity. We propose StoryScope, a pipeline that automatically induces a fine-grained,…
saved by
related reading
- If you let AI do your writing, I will come to your house and kill yousamkriss.substack.com
- How does Pangram work?pangram.substack.com
- How to Tell if Something is AI-Writtenhollisrobbinsanecdotal.substack.com
- Can A.I. Produce Writing That We Actually Want to Read? | The New Yorkernewyorker.com
- Why A.I. Isn’t Going to Make Art | The New Yorkernewyorker.com
- How does Pangram work?substack.com
- 2026 Unslop AI-Written Fiction Contest Resultshyperstitionai.com
- Ghostbuster: Detecting Text Ghostwritten by Large Language Modelsarxiv.org
- The future of literary fiction/‘high art’ in the age of AI : r/TrueLitreddit.com
- Pangram 4 Technical Reportpangram-public.s3.us-east-1.amazonaws.com
- God help me am I about to change my view on AI-assisted writing?rivalvoices.substack.com
- How to Write Like a Human (Without Sounding Like AI)gptzero.me