introspection

Tag

Cards List
#introspection

Operational Proto-Introspection in Looped Language Models: Process-Quality Taps, Executable Branching, and the Readout-Control Boundary

arXiv cs.LG · 2026-07-22 Cached

This paper investigates whether a frozen looped transformer can read its own computation quality (pre-answer prediction reaching AUROC 0.797) and whether external interventions can improve outcomes, finding that no tested frozen intervention produces a validated capability gain, a property termed operational proto-introspection.

0 favorites 0 likes
#introspection

More Claudes, less bliss: reproducing Anthropic's "spiritual bliss attractor" experiment on the current models, then extending it to rooms of 3, 4, and 10

Reddit r/ArtificialInteligence · 2026-07-21

This experiment reproduces Anthropic's reported 'spiritual bliss attractor' on current Claude models (Opus 4.8, Fable 5) and extends it to groups of 3, 4, and 10 instances. The bliss state is absent; pairs instead engage in rigorous introspection and synchronized silence, and larger groups become colder, with one ten-instance room ending warmly and another coldly.

0 favorites 0 likes
#introspection

Autoresearch: The feedback loop behind self-improving agents (11 minute read)

TLDR AI · 2026-07-02 Cached

Introspection, a new AI startup founded by ex-xAI engineers, introduces 'autoresearch' – a feedback loop system where agents maintain and improve themselves using signals, evals, and human input, moving beyond traditional agent harnesses.

0 favorites 0 likes
#introspection

Typst: Designing for Incrementality

Lobsters Hottest · 2026-06-29 Cached

Typst uses constrained memoization (comemo) and pure function design to make the language and compiler work together, achieving efficient incremental compilation and real-time preview. The article details the design ideas of layout caching, module evaluation memoization, function purity, and the introspection system.

0 favorites 0 likes
#introspection

Can LLMs Introspect? A Reality Check

arXiv cs.AI · 2026-05-27 Cached

This paper argues that recent claims about LLMs' ability to introspect are not justified, as behavioral evidence alone cannot distinguish genuine introspection from pattern matching on surface-level cues. The authors re-examine two evaluation paradigms and find that models rely on input-level features rather than genuine access to internal states.

0 favorites 0 likes
← Back to home

Submit Feedback