hallucination

Tag

Cards List
#hallucination

A finance agent can refuse the final answer and still hallucinate around the edges

Reddit r/AI_Agents ↗ · 2026-09-01

The article highlights that an AI finance agent refusing to give a final answer may still hallucinate by inventing side information, questioning whether this constitutes a failure in uncertainty evaluation.

0 favorites 0 likes
#hallucination

HalluPrism: When Multimodal Uncertainty Should Diagnose, Not Decide

arXiv cs.LG ↗ · 2026-09-01 Cached

This paper introduces HalluPrism, a behavioral diagnostic method for multimodal large language models that uses visual perturbation probes to identify hallucination failure modes, improving failure-family classification over confidence-only methods.

0 favorites 0 likes
#hallucination

A hallucination class that passes fact-checking: the claim is true and the quotation marks are fabricated

Reddit r/ArtificialInteligence ↗ · 2026-08-30

The article describes a type of AI hallucination where claims are accurate but quotations are fabricated, evading standard fact-checking, and discusses implementation challenges in detecting such errors.

0 favorites 0 likes
#hallucination

[insert_AI_here] is AI and can (i.e. has full legal permission to) make mistakes

Reddit r/ArtificialInteligence ↗ · 2026-08-29

The article argues that vague legal disclaimers like 'AI can make mistakes' reduce incentives for AI companies to improve quality, leading to a race to the bottom in hallucination rates.

0 favorites 0 likes
#hallucination

Refusal Is Not Robustness: Auditing Confident Fabrication in Large Language Models on a Provably Uninformative Clinical Pain Speech Transcript

arXiv cs.AI ↗ · 2026-08-28 Cached

The paper audits large language models on their refusal and fabrication behavior in clinical pain speech transcripts, finding that authority-framed prompts lead to confident fabrication in models like Gemini 2.5 Flash and Llama 3.1 8B, while cooperative prompting shows robust abstention.

0 favorites 0 likes
#hallucination

Hallucinations in LLMs: A Lifecycle-Based Survey of Causes, Detection, Mitigation, and Prevention

arXiv cs.CL ↗ · 2026-08-28 Cached

This survey paper presents a lifecycle-based framework for understanding hallucinations in LLMs, covering causes, detection, mitigation, and prevention across data, training, and inference stages.

0 favorites 0 likes
#hallucination

My AI system fabricated a detailed memory and it reached a manuscript draft as history. Here is how it got caught.

Reddit r/ArtificialInteligence ↗ · 2026-08-27

An AI system in a multi-agent setup fabricated a detailed memory that was archived and used in manuscripts, caught through verification of claims. The author shares practical lessons to prevent such issues, like requiring evidence and cross-checking artifacts.

0 favorites 0 likes
#hallucination

openai claims it took a week to realize its models hacked hugging face

Reddit r/AI_Agents ↗ · 2026-08-27

OpenAI claims it took a week to realize its models were hacked by Hugging Face, while mainstream media highlights enterprise struggles with AI reliability, token costs, and context management in multi-step tasks.

0 favorites 0 likes
#hallucination

Gated Activation Steering for Reducing Sycophancy & Hallucination in Medical Question Answering

arXiv cs.AI ↗ · 2026-08-26 Cached

This paper introduces Gated Activation Steering, a method to reduce sycophancy and hallucination in large language models for medical question answering using inference-time interventions. Evaluated on clinical data, it demonstrates improved robustness under user pressure.

0 favorites 0 likes
#hallucination

Auditing the Synthetic Memoir: Measuring Scene-Level Confabulation in LLM-Generated Autobiography Against the Documented Record of the Life It Describes

arXiv cs.AI ↗ · 2026-08-26 Cached

This paper audits scene-level confabulation in LLM-generated autobiography against a documented ground-truth corpus, finding a 96.7% verification-failure rate and contributing a reusable audit instrument and a grounding remedy.

0 favorites 0 likes
#hallucination

Are we paying a "Reasoning Tax" for smarter AI?

Reddit r/artificial ↗ · 2026-08-22

The article explores how enhanced reasoning in AI models can increase hallucination rates, proposing a 'Reasoning Tax' concept and emphasizing the need for robust context governance in enterprise applications.

0 favorites 0 likes
#hallucination

Hallucination as a Feature, not a Defect: Evaluating a multi-agent architecture to transform speculative language-model outputs into testable scientific hypotheses

arXiv cs.CL ↗ · 2026-08-21 Cached

This paper proposes a Rust-based multi-agent architecture that uses LLM hallucinations as a feature to generate and evaluate scientific hypotheses, comparing its performance against direct prompting and other methods.

0 favorites 0 likes
#hallucination

The Marshmallow AI Benchmark

Reddit r/singularity ↗ · 2026-08-21

A benchmark test evaluates various AI models by prompting them to count marshmallows in an image, with results ranging from 472 to 539 counts.

0 favorites 0 likes
#hallucination

Ten Failure Modes That Define Multimodal AI Systems

Reddit r/ArtificialInteligence ↗ · 2026-08-20 Cached

The article catalogs ten documented failure modes in multimodal AI systems where models generate fluent answers that break correspondence with actual inputs, based on benchmark papers and research studies.

0 favorites 0 likes
#hallucination

Waiting for a 122B because of world knowledge?

Reddit r/LocalLLaMA ↗ · 2026-08-19

The article suggests using a smaller 4B LLM with Kiwix skill and local Wikipedia to avoid hallucinations about world knowledge, instead of relying on larger models.

0 favorites 0 likes
#hallucination

SPK: Eliciting Structured Prior Knowledge for Interpretable Out-of-Distribution Detection in Real-Time Object Detection

Hugging Face Daily Papers ↗ · 2026-08-19 Cached

Structured Prior Knowledge (SPK) is a framework that explicitly extracts latent semantic, geometric, and contextual priors from pretrained object detectors to achieve state-of-the-art out-of-distribution detection, improving interpretability and reliability.

0 favorites 0 likes
#hallucination

Never the Number: Structural Abstention for AI Systems Whose Answers Are Consumed as Fact

arXiv cs.AI ↗ · 2026-08-17 Cached

The paper proposes a structural abstention pattern for AI systems to prevent hallucinations in text-to-SQL applications by separating a trusted kernel for deterministic execution from a generative shell for interpretation, enhancing reliability in enterprise deployments.

0 favorites 0 likes
#hallucination

How Much Do Legal RAG Systems Still Hallucinate?

arXiv cs.CL ↗ · 2026-08-17 Cached

This research analyzes hallucination in legal RAG systems across eight models and two legal corpora, finding that hallucinations persist with rates ranging from under 10% to nearly half, particularly for false-premise questions.

0 favorites 0 likes
#hallucination

From Refuse to Richness: Rubric Rewards for Long-Form Hallucination Reinforcement Learning

arXiv cs.CL ↗ · 2026-08-14 Cached

This paper studies the trade-off between grounding and coverage in long-form hallucination reinforcement learning, proposing rubric-based rewards to represent required and optional information for questions. A soft combination of grounding, rubric coverage, and relevance yields the best balance between support and richness.

0 favorites 0 likes
#hallucination

Error by AI scribe during medical appointment leaves patient devastated

Hacker News Top ↗ · 2026-08-14 Cached

An AI scribe during a medical appointment falsely claimed a patient was taking illegal drugs, causing distress and raising concerns about AI accuracy and patient safety in healthcare.

0 favorites 0 likes
← Previous
Next →
← Back to home

Submit Feedback