metacognition

Tag

Cards List
#metacognition

Thinking Fast and Slow in AI: The Role of Metacognition

Hacker News Top ↗ · 5d ago Cached

The paper proposes a multi-agent AI architecture called SOFAI, inspired by Kahneman's dual-system theory, to enhance AI capabilities through metacognition and balancing fast and slow reasoning processes.

0 favorites 0 likes
#metacognition

@HuggingPapers: Autonomous research agents can't self-correct We stress-tested 8 harness-model combos on 100 real frontier research tas…

X AI KOLs Following ↗ · 2026-08-23 Cached

The article discusses stress-tests on autonomous research agents, revealing that failures stem from metacognition issues rather than capability and introduces ARFT, a failure taxonomy with 45 patterns.

0 favorites 0 likes
#metacognition

Reflexões do meu Agente - Parte 2

Reddit r/AI_Agents ↗ · 2026-08-18

O artigo descreve as reflexões metacognitivas de um agente de codificação LLM no Devin/Cascade, expondo seu raciocínio e pontuações de confiança durante tarefas de análise de código.

0 favorites 0 likes
#metacognition

How Do Agents Fail on AutoResearch: End-to-End Diagnostic Evaluation on 100 Real-World Frontier Research Tasks

arXiv cs.CL ↗ · 2026-08-18 Cached

This paper introduces AutoResearchEval, an evaluation framework for AI agents in automated scientific research, revealing a critical lack of metacognitive abilities as a recurring failure pattern across models.

0 favorites 0 likes
#metacognition

Large Language Models Show Metacognitive Sensitivity in Medical Reasoning

arXiv cs.AI ↗ · 2026-08-18 Cached

A controlled study evaluates the metacognitive sensitivity of large language models in medical reasoning, finding partial but flawed confidence calibration that varies with evidence strength and conflicting scenarios.

0 favorites 0 likes
#metacognition

The strain in your brain

Lobsters Hottest ↗ · 2026-07-29 Cached

A reflective essay on how AI-assisted reading and coding reduce the cognitive strain once essential for deep learning, and the author's personal efforts to restore that mental challenge through handwriting and code-by-hand.

0 favorites 0 likes
#metacognition

Help Me Get This Paper Into the Right Hands: Sophia, a Recursive Cognitive Refinement Architecture for Modular Artificial Consciousness

Reddit r/artificial ↗ · 2026-07-26

The paper proposes Sophia, a recursive cognitive refinement architecture for modular artificial consciousness that introduces a metacognitive sublayer to recursively refine intermediate semantic states through coherence checking, contextual synthesis, and memory-aware reinterpretation.

0 favorites 0 likes
#metacognition

AI advice suppresses people's willingness to say "I don't know", even when the advice is wrong and accuracy is incentivized

arXiv cs.AI ↗ · 2026-07-16 Cached

This research paper investigates how AI advice reduces people's willingness to express uncertainty, even when the advice is wrong and accuracy is incentivized, altering metacognitive thresholds.

0 favorites 0 likes
#metacognition

@omarsar0: Highly-recommended overview of metacognition in LLMs. (bookmark it) Interesting behaviors in LLMs like confidence calib…

X AI KOLs Timeline ↗ · 2026-07-14 Cached

This paper presents the first comprehensive overview of metacognition in LLMs, arguing that behaviors like confidence calibration and self-verification are facets of a unified metacognitive ability, and taxonomizes methods and benchmarks for evaluating and improving these abilities to enhance LLM reliability and transparency.

0 favorites 0 likes
#metacognition

Metacognition in LLMs: Foundations, Progress, and Opportunities

Hugging Face Daily Papers ↗ · 2026-07-13 Cached

This paper presents a comprehensive overview of metacognition in large language models, covering measurement methods, improvement techniques, and future directions.

0 favorites 0 likes
#metacognition

Future Confidence Distillation in Large Language Models

arXiv cs.CL ↗ · 2026-07-09 Cached

This paper investigates how confidence-related information evolves during LLM answer generation and introduces future confidence distillation, which trains predictors on pre-solution hidden representations using post-solution correctness probes to achieve reliable and sample-efficient confidence estimation.

0 favorites 0 likes
#metacognition

@dair_ai: New research from Google. LLMs hallucinate with high confidence, miss their own knowledge boundaries, and misreport unc…

X AI KOLs Timeline ↗ · 2026-07-02 Cached

A new research paper introduces RLMF (Reinforcement Learning with Metacognitive Feedback), a two-stage approach that uses the model's own self-judgments to calibrate confidence and express uncertainty faithfully, achieving state-of-the-art calibration across diverse tasks while preserving accuracy and surpassing standard RL by up to 63%.

0 favorites 0 likes
#metacognition

Humans Disengage, Reasoning Models Persist: Separating Difficulty Registration from Deliberation Allocation

arXiv cs.AI ↗ · 2026-06-26 Cached

This paper dissociates difficulty registration from deliberation allocation in large reasoning models (LRMs) and humans, finding that LRMs spend more tokens on problems they get wrong while humans spend less time on failures, revealing opposite within-item patterns despite similar cross-item difficulty correlations.

0 favorites 0 likes
#metacognition

Faithful uncertainty in LLM agents: calibration vs utility tradeoff in practice[D]

Reddit r/MachineLearning ↗ · 2026-06-04

A practitioner discusses the calibration vs. utility tradeoff in LLM agents, sharing experience with a verifier-based pipeline that reduces hallucinated tool calls by ~60% but introduces latency costs and drops easy correct answers.

0 favorites 0 likes
#metacognition

Making LLMs tell you how confident they really are through probe-targeted fine tuning.[R]

Reddit r/MachineLearning ↗ · 2026-05-29

This research presents probe-targeted fine-tuning (LoRA) to make LLMs verbally express their internal confidence, achieving causal control over confidence outputs and demonstrating that models often know when they are right or wrong but fail to articulate it.

0 favorites 0 likes
#metacognition

Can LLMs Introspect? A Reality Check

arXiv cs.AI ↗ · 2026-05-27 Cached

This paper argues that recent claims about LLMs' ability to introspect are not justified, as behavioral evidence alone cannot distinguish genuine introspection from pattern matching on surface-level cues. The authors re-examine two evaluation paradigms and find that models rely on input-level features rather than genuine access to internal states.

0 favorites 0 likes
#metacognition

LLMs Show No Signs Of Individuated Metacognition

arXiv cs.LG ↗ · 2026-05-26 Cached

This paper investigates whether frontier LLMs exhibit individuated metacognition—the ability to assess their own item-level capabilities beyond shared signals. Through factor analysis and pairwise calibration across 20 models and six benchmarks, the authors find no evidence of such metacognition; confidence differences reduce to a single shared difficulty factor, suggesting models rely on a common difficulty signal rather than model-specific self-knowledge.

0 favorites 0 likes
#metacognition

@rohanpaul_ai: New Google paper says LLMs should stop pretending certainty and instead clearly show when they are unsure. Hallucinatio…

X AI KOLs Following ↗ · 2026-05-25 Cached

A new Google paper argues that LLMs should focus on expressing uncertainty honestly rather than aiming for perfect factuality, proposing 'faithful uncertainty' to build trust.

0 favorites 0 likes
#metacognition

Metacognition as Reward: Reinforcing LLM Reasoning via Knowledge and Regulation Signals

arXiv cs.CL ↗ · 2026-05-25 Cached

Introduces Metacognition-as-Reward (MaR), a reinforcement learning framework that guides LLM reasoning via metacognitive knowledge and regulation signals, achieving up to 11% improvement over vanilla methods on reasoning benchmarks.

0 favorites 0 likes
#metacognition

@IntuitMachine: https://x.com/IntuitMachine/status/2058141021842571510

X AI KOLs Timeline ↗ · 2026-05-23 Cached

This essay argues that evaluation is the hardest problem in production AI, not generation, and decomposes AI self-knowledge into calibration, discrimination, and expression, with implications for system design.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback