@LangChain: Manual review works at small scale. At millions of agent runs a month, it falls apart, so @Clay relies on LangSmith's o…
Summary
Clay relies on LangSmith's online evaluators to scale manual review for millions of agent runs per month and is testing its insights product to better understand agent behavior.
View Cached Full Text
Cached at: 09/10/26, 04:33 PM
Manual review works at small scale. At millions of agent runs a month, it falls apart, so @Clay relies on LangSmith’s online evaluators and is testing its insights product to understand agent behavior.
@jeffbarg explains. https://t.co/0mjTRiYuBi
Similar Articles
@LangChain: In 13 minutes, @jeffbarg, Vyshu Khota, and Soroush Khadem walk through how Clay scaled agent evals agents at 300M+ runs…
Clay scaled agent evaluations to over 300 million runs per month, covering their four-quadrant eval framework and the challenges of closing the production-to-eval loop.
@LangChain: Improving agents The old way: Manually reading traces, looking for patterns, writing evals, and creating fixes. The bet…
This tweet contrasts the old manual approach to improving AI agents with a new automated method using LangSmith Engine, which cycles through tracing, eval, and fixes.
@LangChain: When your agents take a tumble, LangSmith helps them get back up. LangSmith Evaluation lets you evaluate performance wi…
LangSmith Evaluation helps improve AI agent quality by evaluating performance with real production data.
@LangChain: Healthcare organizations are using agents to transform patient care across workflows. But one persistent bottleneck is …
LangSmith is helping healthcare organizations transform expert reviews into reusable evaluations for AI agents, improving patient care workflows.
@LangChain: Great thread from @AdamRLucek on LangSmith Engine!
Adam Łucek discusses LangSmith Engine, an agent for automating agent development using trace data.