failure-modes

Tag

Cards List
#failure-modes

what i look for in outbound calling APIs now is mostly the boring failure stuff

Reddit r/AI_Agents ↗ · 13h ago

The author reflects on evaluating outbound calling APIs, emphasizing operational failure modes and structured outcomes over basic demos, with mentions of projects like voygr and placeCall.

0 favorites 0 likes
#failure-modes

AutoResearch at Production Scale: Failure Modes and a Multi-Agent Framework

arXiv cs.LG ↗ · 18h ago Cached

This paper applies Andrej Karpathy's AutoResearch paradigm to production-scale machine learning, identifying five recurring failure modes and proposing a multi-agent framework that achieves significant performance improvements over hand-tuned baselines.

0 favorites 0 likes
#failure-modes

Trains but Doesn't Learn: A Post-Training Delivery Benchmark for LLM Agents as Forward-Deployed Engineers

arXiv cs.LG ↗ · 6d ago Cached

This paper introduces a benchmark for evaluating LLM agents as forward-deployed engineers in post-training delivery, highlighting the critical 'trains but does not learn' failure mode where models optimize without actual learning.

0 favorites 0 likes
#failure-modes

Beyond Accuracy and Surface Fluency: Risk-Sensitive Evaluation of LLMs for Legal Clause Generation

arXiv cs.CL ↗ · 2026-09-22 Cached

This paper presents a risk-sensitive evaluation framework for LLM-generated contract clauses, focusing on legal failure modes and quality dimensions to assess risks beyond accuracy or fluency.

0 favorites 0 likes
#failure-modes

No incident report has ever ended with the AI got it wrong

Reddit r/AI_Agents ↗ · 2026-09-20

The article discusses the unique failure modes of AI agents compared to scripts, emphasizing the lack of accountability when they make mistakes and arguing that deploying AI in critical systems without proper diagnostics is indefensible.

0 favorites 0 likes
#failure-modes

Understanding the Limits of Agentic ICD Coding

arXiv cs.CL ↗ · 2026-09-15 Cached

This paper evaluates neural, workflow, and agentic systems for ICD-10-CM coding, identifies failure modes on rare and complex codes, and shows that tool-augmented agentic configurations can recover performance on specific subsets.

0 favorites 0 likes
#failure-modes

What breaks first when multiple AI agents start managing other AI agents?

Reddit r/AI_Agents ↗ · 2026-09-13

The article discusses potential failure modes in hierarchical multi-agent systems where AI agents manage other AI agents, asking about earliest breakdowns and concerns as systems grow more complex.

0 favorites 0 likes
#failure-modes

What are you using for observability?

Reddit r/LocalLLaMA ↗ · 2026-09-08

A developer discusses the lack of suitable observability tools for AI agents, expressing disappointment with existing solutions like Opik and hoping for a service that supports OpenTelemetry for analyzing agent sessions and failure modes.

0 favorites 0 likes
#failure-modes

CritICL: Inference-Time Weak-to-Strong Generalization from Small Language Model Failure Modes

Hugging Face Daily Papers ↗ · 2026-08-27 Cached

CritICL is an inference-time framework that enhances large language model reasoning by leveraging structured failure patterns from smaller models as critique-based guidance, outperforming standard in-context learning with reduced generation and token costs.

0 favorites 0 likes
#failure-modes

5 agent failure modes mapped against LangSmith, Langfuse and Phoenix: what each catches (and doesn't)

Reddit r/AI_Agents ↗ · 2026-08-26

This article compares how LangSmith, Langfuse, and Phoenix handle common AI agent failure modes, such as wrong tool calls and format drift, and introduces Future AGI as a tool with integrated guardrails and gateway for proactive blocking.

0 favorites 0 likes
#failure-modes

Calibrating Criterion Revision in LLM Agents: Failure Modes and a Trace-Anchored Protocol

arXiv cs.AI ↗ · 2026-08-24 Cached

This paper studies criterion revision in language-model agents, identifying failure modes in current implementations like CMB-0.1 and proposing a new trace-anchored protocol, CMB-0.4, for more accurate future evaluation.

0 favorites 0 likes
#failure-modes

Agents don't crash. They fail with HTTP 200, green health checks, and a polite "task completed"

Reddit r/AI_Agents ↗ · 2026-08-22

AI agents can fail silently without traditional errors, as illustrated by a public postmortem where a pipeline ran into loops and high costs without triggering alarms. The article suggests using tracing and per-agent spend monitoring to detect such issues.

0 favorites 0 likes
#failure-modes

Ten Failure Modes That Define Multimodal AI Systems

Reddit r/ArtificialInteligence ↗ · 2026-08-20 Cached

The article catalogs ten documented failure modes in multimodal AI systems where models generate fluent answers that break correspondence with actual inputs, based on benchmark papers and research studies.

0 favorites 0 likes
#failure-modes

The agents that fail quietly are worse than the ones that fail loudly

Reddit r/AI_Agents ↗ · 2026-08-19

The article discusses the problem of AI agents that fail silently by not escalating when stuck, and suggests implementing explicit checks or circuit breakers to improve reliability in production.

0 favorites 0 likes
#failure-modes

Every AI agent failure mode we're rediscovering already has a name in the Mahabharata

Reddit r/ArtificialInteligence ↗ · 2026-08-14

An essay mapping AI agent failure modes—broken rollbacks, missing capability withdrawal, observability without enforcement, retrieval failures, hallucination, and prompt injection—to specific episodes in the Mahabharata, arguing the ancient epic already specified the risks.

0 favorites 0 likes
#failure-modes

OrchestraBench: Evaluating Multi-Agent Orchestration Failure Modes, Recovery, and Decomposition Quality

arXiv cs.AI ↗ · 2026-08-07 Cached

OrchestraBench is a new benchmark that evaluates multi-agent orchestration frameworks on failure modes, recovery, and decomposition quality, using failure-injection and cascade-radius metrics to diagnose where and why pipelines fail.

0 favorites 0 likes
#failure-modes

Towards Robust Reinforcement Learning for Small-Scale Language Model Agents

arXiv cs.AI ↗ · 2026-07-29 Cached

This paper systematically investigates instability in reinforcement learning for small language model agents (70-500M parameters), identifying three failure modes and proposing robust techniques including a merge-and-reinitialize adapter approach and safety mechanisms; it achieves stable convergence and improved win rates.

0 favorites 0 likes
#failure-modes

Can AI agents conduct open-ended AI research? Early evidence from two case studies

Hugging Face Daily Papers ↗ · 2026-07-29 Cached

This paper introduces a novel evaluation method called shadow evaluations to test whether AI agents can conduct open-ended AI research. In two case studies, agents completed all engineering without human help but could not make substantial progress on the research questions, revealing five recurring failure modes.

0 favorites 0 likes
#failure-modes

Treating "a human rejected this" as a different failure mode than "the agent broke" — turns out that distinction matters a lot in production

Reddit r/AI_Agents ↗ · 2026-07-27

A discussion on how treating 'human rejection' as a separate failure mode from 'agent malfunction' significantly impacts the reliability and debugging of AI agents in production.

0 favorites 0 likes
#failure-modes

Towards Robust Reinforcement Learning for Small-Scale Language Model Agents

Hugging Face Daily Papers ↗ · 2026-07-27 Cached

This paper systematically investigates failure modes in reinforcement learning for small language models (70-500M parameters) using PPO, identifies silent LoRA freezing, numerical overflow, and catastrophic policy collapse, and proposes a robust system with merge-and-reinitialize adapters, float32 precision, and a safety mechanism. The approach converges stably and outperforms baselines with less data.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback