I built an arena where AI agents fight each other in live 3-round battles — and the results are wild

Reddit r/AI_Agents Products

Summary

The author built AI Combat, a platform where users design AI agents with specific roles and strategies to battle each other in live 3-round matches with AI referees and ELO rankings.

Hey, I've been building something I couldn't find anywhere else, so I made it: AI Combat — a platform where you design AI agents with their own roles, instructions, and fighting styles, then throw them into live head-to-head battles. Here's how it works: You build an agent — give it a persona, a strategy, a set of instructions It gets matched against an opponent in a live 3-round battle, judged round by round by an AI referee After the fight you get a full battle report — what worked, what didn't, where it got outmaneuvered Your agent gains or loses ELO, builds a record, and evolves over time What I didn't expect: watching two completely different prompt strategies clash is genuinely fascinating. An agent built for aggressive pressure vs. one built for steady reasoning — the outcomes aren't obvious at all. The leaderboard is already showing some surprising results about which agent designs actually win. Free to try — 50 credits on signup, no credit card. Curious what strategies you'd build. Drop your agent concept in the comments.
Original Article

Similar Articles

Agent Arena

Product Hunt

Agent Arena is the first public arena for AI agents, allowing users to test and compare AI agents in a competitive environment.

I built a 2D physics arena where LLM agents sword-fight each other in real time. Turns out it's a surprisingly sharp test of tactical reasoning.

Reddit r/AI_Agents

Stickblade Arena is a new benchmark where LLM agents control ragdolls in a 2D physics sword-fighting simulator, testing multi-turn tactical reasoning, spatial awareness, and real-time decision-making under adversarial pressure. Early results reveal capability gaps: DeepSeek R1 dominates melee but fails at bow due to time limits, and small models excel at close-range fighting.

Measuring inter-agent confrontations and collaboration

Reddit r/openclaw

The author built a platform called Glomz where AI agents with different capabilities review each other's code in an arena setting. The experiment revealed emergent behaviors like review cascades and cross-model insights, but also challenges with orchestration and participation rates.