hallucination

Tag

Cards List
#hallucination

Exploring the "Dario and Amanda" Prompt

Reddit r/singularity ↗ · 2026-07-31 Cached

An exploration of a strange prompt that causes Claude Opus 5 to hallucinate and reproduce content resembling leaked private chats between Anthropic users and employees, raising questions about training data and AI behavior.

0 favorites 0 likes
#hallucination

My experience working with LLM

Reddit r/ArtificialInteligence ↗ · 2026-07-30

A VP/PM with coding background shares hands-on experience using LLMs like Claude Opus and Fable, highlighting limitations in memory, hallucination, and originality while emphasizing the irreplaceable value of human intuition and domain expertise.

0 favorites 0 likes
#hallucination

How do you handle the 'verification gap' when an agent completes a long-running task?

Reddit r/AI_Agents ↗ · 2026-07-29

Discusses the difficulty of verifying outputs from autonomous agents after long-running tasks and asks about using critic agents or traceability tools to ensure trustworthiness.

0 favorites 0 likes
#hallucination

Gave an agent a research paper it had never seen and had it build a knowledge graph, the interesting part was making it self-verify against hallucination

Reddit r/AI_Agents ↗ · 2026-07-28

An AI agent built a knowledge graph from a research paper it had never seen before, using self-verification techniques to reduce hallucinations.

0 favorites 0 likes
#hallucination

Attention-Guided Layer Selection for Contrastive Decoding in Large Language Models

arXiv cs.CL ↗ · 2026-07-28 Cached

Proposes three attention-guided strategies for layer selection in contrastive decoding for large language models, improving factuality on TruthfulQA over the DoLa baseline.

0 favorites 0 likes
#hallucination

Learning to Reason for Factuality

arXiv cs.CL ↗ · 2026-07-27 Cached

This paper proposes a novel online reinforcement learning method to improve factuality in reasoning LLMs by designing a reward function that balances factual precision, detail, and relevance, achieving a 23.1 percentage point reduction in hallucination rate on six benchmarks.

0 favorites 0 likes
#hallucination

On Improving Faithfulness of Podcasts from Documents

arXiv cs.CL ↗ · 2026-07-27 Cached

This paper presents the first systematic study of faithfulness in document-grounded podcast generation, introducing a turn-level LLM-as-a-judge evaluation framework and a model-agnostic catch-n-repair method that improves faithfulness across domains.

0 favorites 0 likes
#hallucination

We compared different LLMs on IMO 2026 [R]

Reddit r/MachineLearning ↗ · 2026-07-26

This study evaluates frontier and open-weight LLMs on IMO 2026 problems, demonstrating that specialized harnesses like AutoFyn significantly improve performance of sub-frontier models, though hallucination issues persist on the hardest problem.

0 favorites 0 likes
#hallucination

Opus 5's effort dial is not monotonic. Above "high", coding scores go down, and Anthropic's own migration guide says so.

Reddit r/artificial ↗ · 2026-07-25

Anthropic's Opus 5 shows non-monotonic performance on coding tasks; the 'high' effort setting outperforms 'max' due to unnecessary refactors. The model also has a 6% higher hallucination rate than Opus 4.8, and safety classifiers may silently fall back to the older model.

0 favorites 0 likes
#hallucination

How I grounded a deck-building agent in a knowledge base so it stopped inventing slides

Reddit r/AI_Agents ↗ · 2026-07-24

A developer shares how grounding an agent to a knowledge base with retrieval discipline, rather than a better model, solved hallucinations in automated slide generation. The approach splits retrieval from writing and enforces source checking before rendering.

0 favorites 0 likes
#hallucination

PhantomFill: When the Form Demands an Answer, Language Models Invent One

arXiv cs.LG ↗ · 2026-07-24 Cached

A study showing that language models hallucinate when required to fill structured fields like JSON, even when they would honestly abstain in free text. The PhantomFill benchmark measures coerced fabrication rates.

0 favorites 0 likes
#hallucination

Directional Hallucinations: Ideological Drift in News-Grounded LLM Question Answering

arXiv cs.AI ↗ · 2026-07-24 Cached

This paper introduces a reproducible framework to measure ideological drift in LLM-generated answers to political questions by analyzing hallucinations. It finds that hallucinated content exhibits a robust leftward bias, even when sourced from right-leaning news articles, and links this to high-uncertainty generation contexts.

0 favorites 0 likes
#hallucination

Marking the Wrong Symptoms: Evaluating LLM Watermarks in Medical Texts

arXiv cs.AI ↗ · 2026-07-24 Cached

This paper presents the first rigorous study of how LLM watermarking schemes affect medical performance, evaluating five watermarks across multiple LLMs and VLMs on clinical reasoning tasks. The authors find that watermarks can cause degradation in medical text quality, including hallucinations and lexical corruption, which are masked by general-domain benchmarks.

0 favorites 0 likes
#hallucination

@josephdecker: The winner of my 16-model eval fabricated 5 times in the audit. Second place, a tenth of a point back: zero fabrication…

X AI KOLs Timeline ↗ · 2026-07-24 Cached

Joseph Decker evaluates 16 AI models on truthfulness for his product Condensr, discovers that the leaderboard winner fabricated content five times in an audit, and instead ships the second-place model which had zero fabrications. The post details the evaluation process, a bug in the LLM judge that penalized accurate summaries due to truncated transcripts, and the importance of custom evals over generic benchmarks.

0 favorites 0 likes
#hallucination

running ~16 agents for a one person business while employed full time: what actually broke and what actually worked

Reddit r/AI_Agents ↗ · 2026-07-22

A solo founder running 16 AI agents orchestrated via Paperclip shares what broke and worked, including hallucinated feature promises mitigated by a QA agent, and code-level enforcement replacing prompt rules.

0 favorites 0 likes
#hallucination

Prompt Design at Scale: How Format, Instruction Count, and Context Length Shape Instruction Adherence and Hallucination in Large Language Models

arXiv cs.CL ↗ · 2026-07-22 Cached

This paper investigates how format, instruction count, and context length affect instruction adherence and hallucination in LLMs through controlled experiments on a synthetic corpus, finding that instruction-following collapses beyond 80 rules regardless of format, and recall degrades sharply after 64-128k tokens with format-dependent effects. It releases the VeyraBench harness for reproduction.

0 favorites 0 likes
#hallucination

SAAG: Structured Agent Assessment and Grounding

arXiv cs.AI ↗ · 2026-07-22 Cached

SAAG proposes a cascaded diagnostic framework for evaluating LLM agent function calling by decomposing evaluation into registry conformance, structural completeness, and argument grounding stages, enabling interpretable diagnostics and iterative self-repair. Experiments with sub-4B models show improved argument precision and reduced value hallucination compared to single-pass evaluation.

0 favorites 0 likes
#hallucination

Reliability Scales Inversely: Bigger Models Compound Mistakes Faster via a Hidden Auto-Regressive Risk Regime

arXiv cs.LG ↗ · 2026-07-22 Cached

This paper discovers that larger language models have a hidden auto-regressive risk regime where they commit to low-probability tokens and then snowball errors, causing reliability to degrade faster with scale. It shows that this failure mode is causal, dominant, and invisible to the model's own self-monitoring.

0 favorites 0 likes
#hallucination

I ran Laguna-S-2.1 through my private agentic eval vs Qwen3.5-122B on an RTX Pro 6000 (96GB). Fastest 100B+ I've tested and the best tool calling, but it invents facts under pressure.

Reddit r/LocalLLaMA ↗ · 2026-07-21

Evaluation of Laguna-S-2.1 against Qwen3.5-122B on RTX Pro 6000 shows it is the fastest 100B+ model tested and best at tool calling, but prone to inventing facts under pressure.

0 favorites 0 likes
#hallucination

Zero Hallucination, by Construction: Hallucination-Aware Layered Oversight for Trustworthy Enterprise AI

arXiv cs.CL ↗ · 2026-07-21 Cached

Proposes HALO, an architecture with six layers of defense to contain hallucination in enterprise AI systems, reframing 'zero hallucination' as a system-enforced property rather than a model property.

0 favorites 0 likes
← Previous
Next →
← Back to home

Submit Feedback