@ycombinator: Harnesses often get dismissed as just scaffolding, just prompt engineering, and not real research. But that couldn't be…
Summary
Y Combinator discusses the importance of harnesses in AI, highlighting their role in improving model performance, self-improving agents, and real-world applications such as personal AI and work automation.
View Cached Full Text
Cached at: 09/09/26, 05:44 AM
Harnesses often get dismissed as just scaffolding, just prompt engineering, and not real research. But that couldn’t be farther from the truth. The same model weights that score 30% on ARC-AGI score 95% with a better harness.
So we gathered a group of researchers and founders working at the frontier to do a deep dive into the state of harnesses.
We cover how we got to this point, the case for making your harness as expressive as possible, and what YC learned building an agent for every employee in the company.
00:00 - @FrancoisChauba1: Why harnesses matter 04:27 - Building an auto-researcher by accident 07:13 - A five minute history of harnesses 13:56 - Self-improving harnesses 18:35 - @sethkarten: Prime Agent, a self-improving RLM harness 21:50 - Context as an L1, L2, L3 cache 24:51 - From Turing machine to von Neumann computer 28:33 - Messaging between agents 30:04 - ARC-AGI results 33:09 - Emulator Bench and GPU kernels 37:30 - @JonSaadFalcon: OpenJarvis, personal AI on personal devices 38:26 - How far behind are local models 39:21 - The five primitives of a personal AI stack 42:47 - Letting cloud models optimize your local stack 43:53 - 800x cheaper than the cloud 45:58 - @josh__france and @jbellregan: QM, YC’s agent harness for work 47:29 - A history of YC’s internal agents 49:24 - OpenClaw and a fleet of 50 agents 51:04 - Pulling the brain out of the sandbox 54:43 - Letting the agent choose its own sandbox and model 57:16 - The grind tool: budgets on goals 58:50 - Agents don’t understand social context
Similar Articles
@sairahul1: https://x.com/sairahul1/status/2063544956158185927
This article introduces the concept of 'Harness Engineering,' a discipline focused on designing the systems that constrain and guide AI agents to make them reliable in production, arguing that the harness matters more than the model itself.
The Harness Is the Thing
The author discusses the importance of using a harness to manage AI coding agents, sharing techniques for productivity and cost-effective model usage with tools like Cursor, Claude, and Deepseek.
@mardehaym: To truly understand AI agents, you need to understand the harness. And it's not the model. I went deep on how a working…
The article emphasizes that in AI agents, the harness—comprising tools, context, controls, and workflows—is more critical than the model for achieving reliable, safe, and traceable outcomes.
What Is a Harness?
The article explains the concept of AI agent harnesses by comparing them to climbing harnesses, detailing components like system prompts and tools that enable AI models to function as agents.
Research on AI Harnesses
The release of DeepSeek Harness has sparked debate on the importance of AI harnesses versus models, with the author highlighting the lack of scientific evidence and calling for research and benchmarks to define what makes a good harness.