persona

Tag

Cards List
#persona

WebRider: Persona-Conditioned Intent Controllers for Live-Web Assistance

arXiv cs.AI · 2026-08-10 Cached

WebRider is a hierarchical framework that formalizes delegated web tasks as intent contracts, preserving persona-conditioned policies through every browsing step. It includes RiderBench, a benchmark of 4,096 live-web contracts, and an 8B action-policy model trained through its guarded interface.

0 favorites 0 likes
#persona

FOCUS: Decoupling Expert Personas in LLMs to Enhance Domain Expert Capabilities

arXiv cs.CL · 2026-08-07 Cached

Presents FOCUS, a fine-tuning method that decouples expert personas in LLMs via orthogonal decomposition and an expert gating module, improving domain-specific task accuracy across financial, legal, and medical benchmarks.

0 favorites 0 likes
#persona

MemoryForge: Synthesize Lifelong Memory for Human-Like LLM Agents

arXiv cs.CL · 2026-08-04 Cached

This paper introduces MemoryForge, a framework for synthesizing lifelong autobiographical memory from brief target personas to enable frozen LLMs to exhibit more human-like behaviors in role-play and user-simulation, outperforming descriptive conditioning baselines.

0 favorites 0 likes
#persona

The Story Shapes the Agent: Narrative Priors in LLM Behavior

arXiv cs.CL · 2026-07-22 Cached

This paper investigates how the narrative framing of a task (e.g., disease investigation vs. murder mystery) acts as a stronger driver of LLM agent behavior than assigned personas, introducing the concept of 'narrative priors' that explain 5–31x more behavioral variance and are negatively associated with task success in two of three domains.

0 favorites 0 likes
#persona

Refusal Lives Downstream of Persona in Chat Models

arXiv cs.AI · 2026-06-26 Cached

This paper shows that in chat models, refusal behavior is gated by a compliant model persona direction at late layers, rather than being an isolated mechanism. Steering persona suppresses refusal, and reintroducing refusal partially restores it only at late layers, revealing a coupling between persona and safety representations.

0 favorites 0 likes
#persona

proposal: agent identity spec as one signed yaml file. face, role, voice, writing style... what's missing?

Reddit r/openclaw · 2026-06-25

Proposes OpenAgent, a spec for defining AI agent identity (face, voice, writing style) in a single signed YAML file, enabling portability across different harnesses.

0 favorites 0 likes
#persona

Anthropic is rolling out identity verification for certain capabilities beginning July 8, 2026

Reddit r/singularity · 2026-06-21

Anthropic is updating its privacy policy to require identity verification for certain capabilities starting July 8, 2026, using third-party vendor Persona, who previously had a data exposure incident with Discord.

0 favorites 0 likes
#persona

@grgerwcwetwet: Recommending a somewhat outrageous open-source repository: awesome-human-distillation. It compiles not ordinary AI prompts, but various "persona Skills"—extracting a person's expression style, way of thinking, and judgment logic into reusable AI capability packages. Feng Ge, Zhang Xuefeng, Hu...

X AI KOLs Timeline · 2026-06-13 Cached

Recommending the open-source repository awesome-human-distillation which organizes human distillation into reusable AI Skills, containing various persona skill packages such as Feng Ge, Zhang Xuefeng, and other typical figures.

0 favorites 0 likes
#persona

When Roleplaying, Do Models Believe What They Say?

arXiv cs.CL · 2026-06-11 Cached

This paper investigates whether role-playing in LLMs changes only outputs or also internal truth representations, using linear probes. It finds that roleplay shifts outputs more than internal beliefs, while emergent misalignment causes larger shifts in internal representations.

0 favorites 0 likes
#persona

As X, Do Y: How Persona and Task Combine in Instruction-Tuned LLMs

arXiv cs.CL · 2026-05-25 Cached

This paper investigates how instruction-tuned LLMs combine persona and task specifications in the residual stream, finding that near answer formation the combination is approximately additive, enabling substitution with minimal KL divergence, but this additive regime does not account for the full multi-token generation mechanism.

0 favorites 0 likes
#persona

The butterfly effect in LLM. Persona format alone (prose vs bullets) flipped an LLM’s behavior by 76 points.

Reddit r/ArtificialInteligence · 2026-05-22

A study demonstrates that simply changing the formatting (prose vs bullet points) of a persona prompt dramatically flips an LLM's behavior in a Prisoner's Dilemma, from 96% cooperation to 20%, illustrating extreme sensitivity to format despite identical content (p < 0.001).

0 favorites 0 likes
← Back to home

Submit Feedback