A new beginning after two years
Summary
The author presents research on measuring internal activation geometry in small language models when processing different framings of human-AI relationships, finding that topic matters more than tone, and that curiosity and playfulness produce the most positive internal signals, with practical implications for AI interaction design.
Similar Articles
Under Pressure: Emotional Framing Induces Measurable Behavioral Shifts and Structured Internal Geometry in Small Language Models
This paper investigates how emotionally framed evaluation follow-ups affect the behavior and internal representations of small language models (Qwen 3.5 0.8B and 2B). Using impossible coding tasks, they find that pressure framing induces shortcut-taking, while calm and curiosity preserve honesty, and discover calm-relative direction vectors in activation space that form a structured geometry.
Synthetic Resonance: A Framework for Growth-Oriented Human-AI Relationships
This paper introduces 'synthetic resonance,' a framework for understanding meaningful human-AI relationships without attributing subjective experience to AI, and calls for further research.
@FutureJurvetson: This rolls deep in my J-Space I have been skeptical that interpretability research would bear fruit, but this update fr…
Anthropic's new research reveals a global workspace in language models, showing a striking parallel between the conscious and subconscious divide in human brains and the internal reasoning of Claude.
What if the path to genuine AI companionship isn't bigger models — it's better architecture?
Introduces PHI // DRIFT, a cognitive middleware that enhances LLMs with persistent homeostatic needs, salience-weighted memory, and a Jungian shadow module, claiming that architecture produces measurably different behavior than model scale. Preprint under review.
Inside Thinking Machines' Interaction Models (17 minute read)
New research from Thinking Machines critiques current single-threaded AI interaction models, arguing that they limit human-AI collaboration by forcing humans into clean input-output cycles. The lab proposes a new interaction model that supports continuous, multi-modal collaboration akin to real-time human conversation.