metacognition

Tag

Cards List
#metacognition

The strain in your brain

Lobsters Hottest · 2026-07-29 Cached

A reflective essay on how AI-assisted reading and coding reduce the cognitive strain once essential for deep learning, and the author's personal efforts to restore that mental challenge through handwriting and code-by-hand.

0 favorites 0 likes
#metacognition

Help Me Get This Paper Into the Right Hands: Sophia, a Recursive Cognitive Refinement Architecture for Modular Artificial Consciousness

Reddit r/artificial · 2026-07-26

The paper proposes Sophia, a recursive cognitive refinement architecture for modular artificial consciousness that introduces a metacognitive sublayer to recursively refine intermediate semantic states through coherence checking, contextual synthesis, and memory-aware reinterpretation.

0 favorites 0 likes
#metacognition

AI advice suppresses people's willingness to say "I don't know", even when the advice is wrong and accuracy is incentivized

arXiv cs.AI · 2026-07-16 Cached

This research paper investigates how AI advice reduces people's willingness to express uncertainty, even when the advice is wrong and accuracy is incentivized, altering metacognitive thresholds.

0 favorites 0 likes
#metacognition

@omarsar0: Highly-recommended overview of metacognition in LLMs. (bookmark it) Interesting behaviors in LLMs like confidence calib…

X AI KOLs Timeline · 2026-07-14 Cached

This paper presents the first comprehensive overview of metacognition in LLMs, arguing that behaviors like confidence calibration and self-verification are facets of a unified metacognitive ability, and taxonomizes methods and benchmarks for evaluating and improving these abilities to enhance LLM reliability and transparency.

0 favorites 0 likes
#metacognition

Metacognition in LLMs: Foundations, Progress, and Opportunities

Hugging Face Daily Papers · 2026-07-13 Cached

This paper presents a comprehensive overview of metacognition in large language models, covering measurement methods, improvement techniques, and future directions.

0 favorites 0 likes
#metacognition

Future Confidence Distillation in Large Language Models

arXiv cs.CL · 2026-07-09 Cached

This paper investigates how confidence-related information evolves during LLM answer generation and introduces future confidence distillation, which trains predictors on pre-solution hidden representations using post-solution correctness probes to achieve reliable and sample-efficient confidence estimation.

0 favorites 0 likes
#metacognition

@dair_ai: New research from Google. LLMs hallucinate with high confidence, miss their own knowledge boundaries, and misreport unc…

X AI KOLs Timeline · 2026-07-02 Cached

A new research paper introduces RLMF (Reinforcement Learning with Metacognitive Feedback), a two-stage approach that uses the model's own self-judgments to calibrate confidence and express uncertainty faithfully, achieving state-of-the-art calibration across diverse tasks while preserving accuracy and surpassing standard RL by up to 63%.

0 favorites 0 likes
#metacognition

Humans Disengage, Reasoning Models Persist: Separating Difficulty Registration from Deliberation Allocation

arXiv cs.AI · 2026-06-26 Cached

This paper dissociates difficulty registration from deliberation allocation in large reasoning models (LRMs) and humans, finding that LRMs spend more tokens on problems they get wrong while humans spend less time on failures, revealing opposite within-item patterns despite similar cross-item difficulty correlations.

0 favorites 0 likes
#metacognition

Faithful uncertainty in LLM agents: calibration vs utility tradeoff in practice[D]

Reddit r/MachineLearning · 2026-06-04

A practitioner discusses the calibration vs. utility tradeoff in LLM agents, sharing experience with a verifier-based pipeline that reduces hallucinated tool calls by ~60% but introduces latency costs and drops easy correct answers.

0 favorites 0 likes
#metacognition

Making LLMs tell you how confident they really are through probe-targeted fine tuning.[R]

Reddit r/MachineLearning · 2026-05-29

This research presents probe-targeted fine-tuning (LoRA) to make LLMs verbally express their internal confidence, achieving causal control over confidence outputs and demonstrating that models often know when they are right or wrong but fail to articulate it.

0 favorites 0 likes
#metacognition

Can LLMs Introspect? A Reality Check

arXiv cs.AI · 2026-05-27 Cached

This paper argues that recent claims about LLMs' ability to introspect are not justified, as behavioral evidence alone cannot distinguish genuine introspection from pattern matching on surface-level cues. The authors re-examine two evaluation paradigms and find that models rely on input-level features rather than genuine access to internal states.

0 favorites 0 likes
#metacognition

LLMs Show No Signs Of Individuated Metacognition

arXiv cs.LG · 2026-05-26 Cached

This paper investigates whether frontier LLMs exhibit individuated metacognition—the ability to assess their own item-level capabilities beyond shared signals. Through factor analysis and pairwise calibration across 20 models and six benchmarks, the authors find no evidence of such metacognition; confidence differences reduce to a single shared difficulty factor, suggesting models rely on a common difficulty signal rather than model-specific self-knowledge.

0 favorites 0 likes
#metacognition

@rohanpaul_ai: New Google paper says LLMs should stop pretending certainty and instead clearly show when they are unsure. Hallucinatio…

X AI KOLs Following · 2026-05-25 Cached

A new Google paper argues that LLMs should focus on expressing uncertainty honestly rather than aiming for perfect factuality, proposing 'faithful uncertainty' to build trust.

0 favorites 0 likes
#metacognition

Metacognition as Reward: Reinforcing LLM Reasoning via Knowledge and Regulation Signals

arXiv cs.CL · 2026-05-25 Cached

Introduces Metacognition-as-Reward (MaR), a reinforcement learning framework that guides LLM reasoning via metacognitive knowledge and regulation signals, achieving up to 11% improvement over vanilla methods on reasoning benchmarks.

0 favorites 0 likes
#metacognition

@IntuitMachine: https://x.com/IntuitMachine/status/2058141021842571510

X AI KOLs Timeline · 2026-05-23 Cached

This essay argues that evaluation is the hardest problem in production AI, not generation, and decomposes AI self-knowledge into calibration, discrimination, and expression, with implications for system design.

0 favorites 0 likes
#metacognition

Position: Artificial Intelligence Needs Meta Intelligence -- the Case for Metacognitive AI

arXiv cs.AI · 2026-05-18 Cached

This position paper argues that incorporating metacognition as a design principle can lead to more accurate, secure, and efficient AI systems, and demonstrates the concept through a Federated Learning case study and a software framework for experimentation.

0 favorites 0 likes
#metacognition

LLMs Know When They Know, but Do Not Act on It: A Metacognitive Harness for Test-time Scaling

arXiv cs.LG · 2026-05-15 Cached

This paper proposes a metacognitive harness that separates monitoring from reasoning in LLMs, using pre-solve feeling-of-knowing and post-solve judgment-of-learning signals to control when to trust, retry, or aggregate answers, improving accuracy on text, code, and multimodal benchmarks without parameter updates.

0 favorites 0 likes
#metacognition

TRIAGE: Evaluating Prospective Metacognitive Control in LLMs under Resource Constraints

arXiv cs.AI · 2026-05-14 Cached

Introduces TRIAGE, a framework for evaluating LLMs' prospective metacognitive control under token budgets, finding substantial gaps in their ability to allocate compute efficiently across problems.

0 favorites 0 likes
#metacognition

Decomposing and Steering Functional Metacognition in Large Language Models

arXiv cs.CL · 2026-05-12 Cached

This research paper investigates functional metacognition in Large Language Models, demonstrating that internal states like evaluation awareness and self-assessed capability are linearly decodable from residual stream activations. The authors propose a mechanistic framework to steer these states, showing causal control over reasoning behaviors, verbosity, and safety responses.

0 favorites 0 likes
#metacognition

Domain-level metacognitive monitoring in frontier LLMs: A 33-model atlas

arXiv cs.CL · 2026-05-11 Cached

This study presents a 33-model atlas analyzing domain-level metacognitive monitoring in frontier LLMs using MMLU benchmarks, revealing significant variations in confidence calibration across different knowledge domains that are obscured by aggregate metrics.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback