Tag
This paper shows that repetition effects in language models depend on readout position: adjacent repetition boosts target probability, while displaced repetition produces an inverted-U curve. The finding challenges assumptions in cloze-style probing and is validated across multiple models and languages.
This arXiv paper proposes capability-sustaining emotional dialogue (CSED) as a longitudinal research paradigm for emotional support systems, arguing that current approaches focus on immediate relief and neglect long-term user capabilities. A literature audit shows most systems overlook longitudinal outcomes, suggesting a new agenda for data, models, evaluation, and governance.
A fine-grained study of narrative features in web-scale LLM pretraining data, introducing NarraBERT and NarraDolma to measure narrative patterns and their distribution across sources.
This paper introduces a method for detecting hallucinations in large language models by leveraging the confidence of the first generated token, requiring only a single decode step.