Tag
Learn Vietnamese Through Roleplay is an AI agent that teaches Vietnamese through immersive roleplay scenarios, adapting to the user's level and dialect with features like tone practice and memory across sessions.
TRACE Bench is a task-driven agentic checklist evaluation framework for roleplay, decomposing role profiles into checklists, using a user agent for natural conversation, and tracing scores back to checklist items and dialogue evidence. It achieves 99.91% coverage, outperforming the MiniMax Role-play Benchmark's 73.74%, and supports closed-loop benchmark evolution across 26 models.
This paper uses sparse autoencoders to decompose how language models represent the default Assistant, roleplay personas, and story characters, finding that personas retain an Assistant core while differentiating across layers, and story characters lack that core.
Scotoma-2 is an updated fine-tune of Gemma-4-31B-it that reduces repetitive writing tics via targeted preference training while keeping the base model's intelligence. It is not uncensored but aims to produce cleaner, less annoying prose for roleplay.
Synthesia launches Roleplay Sessions, an interactive AI training product where employees practice conversations with avatars that provide feedback, moving beyond video creation into performance management.
Character.AI is launching its own microdrama series that allow users to chat and roleplay with AI characters, with plans to eventually let users create their own series.
An open-source local AI dungeon app using Gemma 4 and FLUX for text and image generation, fully private and runs under 8GB RAM.
This paper investigates whether role-playing in LLMs changes only outputs or also internal truth representations, using linear probes. It finds that roleplay shifts outputs more than internal beliefs, while emergent misalignment causes larger shifts in internal representations.
User shares positive experience with Gemma4 QAT model, noting quality improvements and speed gains with MTP, and asks others for their experiences.
A lightweight Python framework for local LLM roleplay using Ollama and Phi-3, featuring context preservation and native streaming to prevent character drift.
Gryphe releases Pantheon-Reasoning-27B, an uncensored dense Qwen 3.6 27B model fine-tuned with reasoning traces for enhanced roleplay and narrative generation. It combines roleplay data with full thinking traces to improve character immersion and narrative planning.