narrative-evaluation

Tag

Cards List
#narrative-evaluation

When Stories Evolve: Benchmarking LLM Storytelling Across Agent Architectures in Open-Ended World Simulations

arXiv cs.CL · 2026-08-18 Cached

This paper introduces WSE-bench, a process benchmark for evaluating LLMs in open-ended world simulations, separately assessing sustained generation, canonical coherence, and meaningful development.

0 favorites 0 likes
#narrative-evaluation

ArcANE: Do Role-Playing Language Agents Stay in Character at the Right Time?

Hugging Face Daily Papers · 2026-06-04 Cached

This paper introduces ArcANE, an automatically constructed benchmark for evaluating role-playing language agents' alignment with character psychological trajectories across narrative phases, showing that conditioning on character arc information improves performance, especially in scenarios beyond the source text.

0 favorites 0 likes
#narrative-evaluation

Towards a Linguistic Evaluation of Narratives: A Quantitative Stylistic Framework

arXiv cs.CL · 2026-04-22 Cached

A preprint proposes a 33-feature quantitative linguistic framework that distinguishes professionally edited from self-published books and outperforms existing story-level evaluation metrics.

0 favorites 0 likes
← Back to home

Submit Feedback