Same task in github-copilot, pi, claude-code, and opencode with Qwen3.6 27B
Summary
The author tests multiple coding agent harnesses (GitHub Copilot, Pi, Claude Code, OpenCode) using the same Qwen3.6 27B model, finding that harness design significantly impacts performance, with OpenCode excelling at web searches and web development, and GitHub Copilot struggling with file editing tools.
Similar Articles
favorite Agentic Coding Harness
The author compares several agentic coding harnesses (Codex CLI, Claude Code, Gemini CLI, OpenCode, Pi) and finds Pi the leanest and best for local models, praising its simplicity and compatibility with Qwen 27B-MXFP8.
Harness showdown: Claude Code vs OpenCode vs Pi with DeepSeek V4 Flash
A comparison article pitting AI coding tools Claude Code, OpenCode, and Pi against DeepSeek V4 Flash in a harness showdown.
I used local Qwen 27b to build a harness and replace OpenCode
The author built an open-source harness for running local LLMs using Qwen 27b, featuring just-in-time code review, sub-agents, voice dictation, and more, and shares development insights.
Been running Qwen3.6-27B through a 3-critic harness. The harness matters more than I thought
Reports on running Qwen3.6-27B (8-bit) through a 3-critic coding harness, finding the harness effectively catches errors and makes final output quality comparable to frontier models, with a proposed workflow of frontier for planning and Qwen for execution.
@MiaAI_lab: Qwopus 3.6-27b Coder I had a lot of requests to test it, so I did. I ran the same tests I’ve done on other models. It s…
MiaAI Lab tested Qwopus 3.6-27b Coder and found it underperformed compared to Qwen 3.6 27b and 35b in tool-calling and code generation, with broken HTML demos.