coding-agent

Tag

Cards List
#coding-agent

Can a coding agent use 57–85% less fresh model traffic without losing task success? I open-sourced my experiment

Reddit r/AI_Agents · 17h ago

The author open-sources an execution and context layer for coding agents that cuts fresh model traffic by 57-85% while preserving task success in paired smoke tests on GPT-5.6 and Claude Opus 5, and seeks independent evaluation and sponsorship.

0 favorites 0 likes
#coding-agent

my coding agent now deploys its own changes to a sandbox and tests them before i merge

Reddit r/AI_Agents · 17h ago

The author describes using Mastra's new preview deployment feature to let their coding agent automatically deploy changes to a sandbox, test them via API and UI, and then open a PR, closing the verification gap.

0 favorites 0 likes
#coding-agent

Auto mode is now the default in Claude Code for Pro, Max, and Team plans

Simon Willison's Blog · yesterday Cached

Anthropic is making auto mode the default in Claude Code for Pro, Max, and Team plans starting August 14, backed by evals claiming auto mode blocks 89% of harmful actions and resists all tested indirect prompt injection attacks.

0 favorites 0 likes
#coding-agent

Claude Code in 9 lines python

Reddit r/LocalLLaMA · yesterday

A developer shares a 9-line Python implementation of a minimal coding agent similar to Claude Code or Codex, using only the standard library and compatible with any OpenAI Responses API, with code on GitHub.

0 favorites 0 likes
#coding-agent

people will compare the model. the more useful part might be how it runs

Reddit r/AI_Agents · 2d ago

Meta released Muse Code beta with Muse Spark 1.2, positioning it as a third coding-agent option next to Claude Code and Codex. Its key differentiators are parallel sub-agents in isolated worktrees, persistent background agents, and local event logging for crash recovery, making it suited for long-running tasks on large repos.

0 favorites 0 likes
#coding-agent

@sashimikun_void: Prime-Bun: https://github.com/sng-asyncfunc/prime-bun… An ultra-fast Prime Agent fork designed to orchestrate concurren…

X AI KOLs Following · 3d ago Cached

Prime Bun is a Bun-native fork of Prime Agent that replaces the Python notebook with a JavaScript/TypeScript runtime, dramatically reducing latency for long-running coding tasks.

0 favorites 0 likes
#coding-agent

@sunmer575399: Stumbled upon a god-tier open-source project, Cline, with 65.6k stars. It really boosts coding efficiency. One engine powers the SDK, CLI, VS Code, and the entire JetBrains suite. Click twice in the editor, and it reads code, creates files, runs commands, and after making changes, waits for your nod before proceeding. Run full-auto in the terminal...

X AI KOLs Timeline · 3d ago Cached

Introducing the open-source AI coding agent tool Cline, supporting the SDK, CLI, VS Code, and the JetBrains suite. It can automatically read code, create files, and run commands in the IDE and terminal, and supports kanban-based parallel multi-agent workflows and CI/CD integration.

0 favorites 0 likes
#coding-agent

Prime Agent - a new coding harness surpassing Codex/CC/PI

Reddit r/LocalLLaMA · 4d ago

Prime Agent is an open-source coding and research harness that outperforms proprietary harnesses, scoring 95.5% on ARC-AGI-3 and improving models across benchmarks.

0 favorites 0 likes
#coding-agent

@LinusEkenstam: Just before bed time. Let me sleep. plz 95.5% on ARC-AGI-3 (’huge if true”)

X AI KOLs Timeline · 4d ago Cached

Linus Ekenstam highlights Prime Intellect's release of Prime Agent, a self-improving harness for coding and long-running autonomous tasks, reportedly scoring 95.5% on ARC-AGI-3, above the human baseline.

0 favorites 0 likes
#coding-agent

Prime Agent: A self-improving RLM agent

Hacker News Top · 4d ago Cached

Prime Intellect launches Prime Agent, a fully open-source self-improving coding harness built around Recursive Language Model (RLM) and Continual Harness abstractions, enabling persistent sub-agents and dynamic tooling via a REPL-based interface.

0 favorites 0 likes
#coding-agent

Muse Code and Muse Spark 1.2

Hacker News Top · 4d ago Cached

Meta releases Muse Code, a terminal coding agent, and Muse Spark 1.2, an upgraded coding-focused model with improved code generation, debugging, and long-horizon task handling.

0 favorites 0 likes
#coding-agent

@yibie: Databricks tested various coding tools with real tasks from their own team—conclusion: the same model called from different harnesses has a per-task cost difference of more than 2x, while quality is the same. "Pi: Minimalism and High Performance" Pi's minimalism is its advantage AI makes code cheaper, and as a result many companies build bigger...

X AI KOLs Timeline · 4d ago Cached

Databricks' benchmark shows that the same model invoked through different harnesses has a cost difference of more than 2x, while Pi, as a minimalist coding harness, delivers high performance at low cost; Shopify also used Pi to extend Autoresearch and improve efficiency.

0 favorites 0 likes
#coding-agent

Pi's Minimalism Is Its Advantage

Hacker News Top · 5d ago Cached

Earendil's Pi coding harness demonstrates that minimalist design outperforms complex alternatives in cost and performance, citing Databricks benchmarks and a Shopify case study as evidence.

0 favorites 0 likes
#coding-agent

My co-founders and I are launching a coding agent with a twist: Unlimited usage. How stupid are we?

Reddit r/artificial · 5d ago

A startup announces a coding agent with unlimited usage at a flat rate, using domain-specific sub-agents to maximize cost efficiency. Early access signups begin soon.

0 favorites 0 likes
#coding-agent

The Warp Agent CLI

Hacker News Top · 5d ago Cached

Warp has launched the Warp Agent CLI, a standalone multi-model coding agent for any terminal, featuring built-in model routing, persistent sessions, remote agents, and native muxing of agent sessions based on Warp's terminal infrastructure.

0 favorites 0 likes
#coding-agent

@pidotdev: Read the full blog post here

X AI KOLs Following · 5d ago Cached

A blog post highlighting Pi, a minimal coding agent harness, arguing that its simplicity yields better performance and lower cost compared to more complex tools, supported by Databricks benchmark results and Shopify's Pi Autoresearch case study.

0 favorites 0 likes
#coding-agent

An agent just coded for 10 days with nobody watching. Qwen 3.8 max

Reddit r/AI_Agents · 6d ago

Alibaba's Qwen agent autonomously coded for over 10 days in an empty repo, filing issues, writing code, running tests, fixing failures, and merging. It still required some feedback, but demonstrates a self-correcting autonomous loop.

0 favorites 0 likes
#coding-agent

Opus 5 vs Opus 4.8 vs GPT-5.6 Sol, tested for free. Model choice was never my problem.

Reddit r/AI_Agents · 6d ago

A solo developer tests Opus 5, Opus 4.8, GPT-5.6 Sol and Kimi K3 via a multi-model router with free credit, discovering that evaluation budgets and input preprocessing matter more than raw model choice.

0 favorites 0 likes
#coding-agent

@ashpreetbedi: Introducing the Pro Agent Builders series. Inside information on how experts are building agents. First post: automatin…

X AI KOLs Timeline · 6d ago Cached

Introduces the Pro Agent Builders series, focusing on recursive auto-improvement (RAI) where a coding agent runs hundreds of probes overnight to automatically improve another agent's performance.

0 favorites 0 likes
#coding-agent

Why is your coding agent idle for the eight hours you are asleep?

Reddit r/AI_Agents · 6d ago

The author describes a workflow for running an AI coding agent overnight with a single goal, constraints, and a mandatory morning report, to maximize idle compute time.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback