failure-modes

Tag

Cards List
#failure-modes

OrchestraBench: Evaluating Multi-Agent Orchestration Failure Modes, Recovery, and Decomposition Quality

arXiv cs.AI · 2026-08-07 Cached

OrchestraBench is a new benchmark that evaluates multi-agent orchestration frameworks on failure modes, recovery, and decomposition quality, using failure-injection and cascade-radius metrics to diagnose where and why pipelines fail.

0 favorites 0 likes
#failure-modes

Towards Robust Reinforcement Learning for Small-Scale Language Model Agents

arXiv cs.AI · 2026-07-29 Cached

This paper systematically investigates instability in reinforcement learning for small language model agents (70-500M parameters), identifying three failure modes and proposing robust techniques including a merge-and-reinitialize adapter approach and safety mechanisms; it achieves stable convergence and improved win rates.

0 favorites 0 likes
#failure-modes

Can AI agents conduct open-ended AI research? Early evidence from two case studies

Hugging Face Daily Papers · 2026-07-29 Cached

This paper introduces a novel evaluation method called shadow evaluations to test whether AI agents can conduct open-ended AI research. In two case studies, agents completed all engineering without human help but could not make substantial progress on the research questions, revealing five recurring failure modes.

0 favorites 0 likes
#failure-modes

Treating "a human rejected this" as a different failure mode than "the agent broke" — turns out that distinction matters a lot in production

Reddit r/AI_Agents · 2026-07-27

A discussion on how treating 'human rejection' as a separate failure mode from 'agent malfunction' significantly impacts the reliability and debugging of AI agents in production.

0 favorites 0 likes
#failure-modes

Towards Robust Reinforcement Learning for Small-Scale Language Model Agents

Hugging Face Daily Papers · 2026-07-27 Cached

This paper systematically investigates failure modes in reinforcement learning for small language models (70-500M parameters) using PPO, identifies silent LoRA freezing, numerical overflow, and catastrophic policy collapse, and proposes a robust system with merge-and-reinitialize adapters, float32 precision, and a safety mechanism. The approach converges stably and outperforms baselines with less data.

0 favorites 0 likes
#failure-modes

after a year of shipping with AI agents, here's what they still reliably get wrong

Reddit r/AI_Agents · 2026-07-24

A developer shares consistent failure modes of AI agents after a year of shipping code with them, including confidently wrong code, inability to maintain cross-file architecture, lacking pushback on bad decisions, and security edge case issues.

0 favorites 0 likes
#failure-modes

Gartner thinks 40% of agentic AI projects get canceled by 2027. Building one right now, I believe it.

Reddit r/AI_Agents · 2026-07-23

The author reflects on Gartner's prediction that 40% of agentic AI projects will be canceled by 2027, emphasizing that the real failure is not model incompetence but quiet failures in production due to bad data or API issues, and that most teams measure single task completion rather than reliability over hundreds of runs.

0 favorites 0 likes
#failure-modes

AI From the Trenches: Why Its Brilliance and Failures Share the Same Root

Reddit r/artificial · 2026-07-22

The author shares two years of experience building a platform with AI, identifying six recurring failure modes (Band-Aid, Assumption, Drift, Hallucination, Lack of Common Sense, Path of Least Resistance) and argues that even as models improve, these failure modes persist, becoming harder to detect.

0 favorites 0 likes
#failure-modes

@BhavinJawade: 𝗢𝗻-𝗽𝗼𝗹𝗶𝗰𝘆 𝗱𝗶𝘀𝘁𝗶𝗹𝗹𝗮𝘁𝗶𝗼𝗻 𝗶𝘀𝗻'𝘁 𝗮 𝗳𝗿𝗲𝗲-𝗹𝘂𝗻𝗰𝗵 On-policy distillation has become a default…

X AI KOLs Timeline · 2026-07-21 Cached

Bhavin Jawade discusses several failure modes of on-policy distillation for training large language models, including early mistakes becoming uncorrectable, stronger teachers being worse, privileged information conditioning failing to transfer, and thinking collapse from dense supervision.

0 favorites 0 likes
#failure-modes

@BhavinJawade: I am surveying papers that discuss and explain the failure modes of on-policy distillation and its variants. Will be sh…

X AI KOLs Timeline · 2026-07-20 Cached

BhavinJawade surveys papers on failure modes of on-policy distillation and its variants, listing several recent arxiv papers including 'The Many Faces of On-Policy Distillation' and others.

0 favorites 0 likes
#failure-modes

Silent Failures in Quantized LLM Reasoning: A Taxonomy-Based Analysis of Hollow Convergence and Failure Mode Shifts

arXiv cs.CL · 2026-07-14 Cached

This paper demonstrates that post-training quantization can silently alter how large language models reason, even when task accuracy is preserved, through a taxonomy-based analysis of 30,000 chain-of-thought outputs across multiple models and benchmarks.

0 favorites 0 likes
#failure-modes

The boring failure modes of paid AI agents are more interesting than the demos

Reddit r/AI_Agents · 2026-07-12

The author discusses the practical failure modes of AI agents that use paid tools, such as cost unawareness, double-spends, and the need for human approval, suggesting that agent payments should be treated as a separate execution layer.

0 favorites 0 likes
#failure-modes

@LiorOnAI: An open-source fix for one of the most common reasoning model failure modes. One of the biggest AI trends this year isn…

X AI KOLs Timeline · 2026-07-07 Cached

Liquid AI releases Antidoom, an open-source method that fine-tunes reasoning models to break repetitive token loops (doom loops), reducing failure rates from ~23% to 1% on Qwen3.5-4B without retraining or RL.

0 favorites 0 likes
#failure-modes

Auditing the Audit: Five Failure Modes in Benchmark-Validity Audits

arXiv cs.LG · 2026-07-07 Cached

This paper identifies five failure modes in perturbation-based benchmark-validity audits used for AI governance, demonstrating that implementation details can silently manufacture conclusions. It proposes a due-diligence gate to improve the reliability of evaluation evidence.

0 favorites 0 likes
#failure-modes

@Sprytixl: ANTHROPIC LEAD ENGINEER MAKING $2.3M/YEAR JUST LEAKED A 12-PAGE DOCUMENT - AND GOT FIRED 15 MINUTES AFTER PUBLISHING mo…

X AI KOLs Timeline · 2026-07-05 Cached

An Anthropic lead engineer leaked a 12-page document detailing five common failure modes in agentic loops and was fired shortly after. The thread summarizes the key failure types including blind, tangled, nodding, amnesiac, and manual loops.

0 favorites 0 likes
#failure-modes

The next failure mode in AI agents will not look like one bad prompt.

Reddit r/AI_Agents · 2026-07-03

This article discusses the next expected failure mode in AI agents, which will likely be more complex than a single bad prompt.

0 favorites 0 likes
#failure-modes

What killed your agent after its first few weeks in production?

Reddit r/AI_Agents · 2026-06-30

This article explores common reasons why AI agents fail shortly after being deployed in production, highlighting pitfalls and lessons learned.

0 favorites 0 likes
#failure-modes

@cyrilXBT: https://x.com/cyrilXBT/status/2070690243880116242

X AI KOLs Timeline · 2026-06-27 Cached

A practical guide explaining why naive multi-agent systems fail and how to build coordinated AI agent teams using Builder, Judge, and Manager roles with clear handoffs and verification.

0 favorites 0 likes
#failure-modes

What are the most common failure modes of AI agents in enterprise environments?

Reddit r/AI_Agents · 2026-06-15

Discusses common failure modes of AI agents in enterprise environments, such as over-reliance on long-term memory and stateless tool gating leading to security risks.

0 favorites 0 likes
#failure-modes

The worst coding agent failure is when it says “done” too early

Reddit r/AI_Agents · 2026-06-13

The article highlights a common failure mode in coding agents where they report tasks as 'done' while leaving hidden issues like insufficient tests, missed edge cases, and introduced bugs, creating a trust problem for developers.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback