@XAMTO_AI: A veteran engineer packed decades of practical engineering experience into this open-source project, which shot to #1 on GitHub trending, amassing 124k stars. The author, a former Vercel engineer who participated in early Next.js development, compiled 16 practical techniques for collaborating with Claude, installable with a single command. The most impressive…
Summary
Former Vercel engineer Matt Pocock open-sourced a project called 'skills' that provides 16 practical tips for collaborating with AI coding agents like Claude, including 'Grill Me' and red-green test cycles. Aimed at solving common AI development issues, it has garnered 124k stars.
View Cached Full Text
Cached at: 06/11/26, 02:08 PM
A veteran developer packed decades of hands-on engineering experience into this open-source project, which shot straight to #1 on the GitHub trending list with 124k stars. The author is a former Vercel engineer who worked on early Next.js development, and has compiled 16 practical tips for collaborating with Claude, installable with a single command. The most impressive is the “Grim” skill — it flips the script and makes the AI interrogate you, walking through a decision tree to confirm requirements one by one, making sure it understands exactly what you want before taking action. With just 42 words, it’s been called “the prompt with the highest token return rate.” There’s also a skill specifically for code that won’t run: it forces the AI to first write a test that must fail, then write the minimal code to pass it — the classic red-green cycle, leaving the AI no chance to cut corners. Link: https://github.com/mattpocock/skills…
mattpocock/skills Source: https://github.com/mattpocock/skills
Skills For Real Engineers
skills.sh (https://skills.sh/mattpocock/skills)
My agent skills that I use every day to do real engineering - not vibe coding.
Developing real applications is hard. Approaches like GSD, BMAD, and Spec-Kit try to help by owning the process. But while doing so, they take away your control and make bugs in the process hard to resolve.
These skills are designed to be small, easy to adapt, and composable. They work with any model. They’re based on decades of engineering experience.
Hack around with them. Make them your own. Enjoy.
If you want to keep up with changes to these skills, and any new ones I create, you can join ~60,000 other devs on my newsletter: Sign Up To The Newsletter
Quickstart (30-second setup)
- Run the skills.sh installer:
bash npx skills@latest add mattpocock/skills - Pick the skills you want, and which coding agents you want to install them on.
Make sure you select
/setup-matt-pocock-skills. - Run
/setup-matt-pocock-skillsin your agent. It will:- Ask you which issue tracker you want to use (GitHub, Linear, or local files)
- Ask you what labels you apply to tickets when you triage them (
/triageuses labels) - Ask you where you want to save any docs we create
- Bam - you’re ready to go.
Why These Skills Exist
I built these skills as a way to fix common failure modes I see with Claude Code, Codex, and other coding agents.
#1: The Agent Didn’t Do What I Want
“No-one knows exactly what they want”
David Thomas & Andrew Hunt, The Pragmatic Programmer
The Problem. The most common failure mode in software development is misalignment. You think the dev knows what you want. Then you see what they’ve built - and you realize it didn’t understand you at all. This is just the same in the AI age. There is a communication gap between you and the agent. The fix for this is a grilling session - getting the agent to ask you detailed questions about what you’re building.
The Fix is to use:
/grill-me- for non-code uses/grill-with-docs- same as/grill-me, but adds more goodies (see below)
These are my most popular skills. They help you align with the agent before you get started, and think deeply about the change you’re making. Use them every time you want to make a change.
#2: The Agent Is Way Too Verbose
With a ubiquitous language, conversations among developers and expressions of the code are all derived from the same domain model.
Eric Evans, Domain-Driven-Design
The Problem: At the start of a project, devs and the people they’re building the software for (the domain experts) are usually speaking different languages. I felt the same tension with my agents. Agents are usually dropped into a project and asked to figure out the jargon as they go. So they use 20 words where 1 will do.
The Fix for this is a shared language. It’s a document that helps agents decode the jargon used in the project.
Example
Here’s an example CONTEXT.md (https://github.com/mattpocock/course-video-manager/blob/076a5a7a182db0fe1e62971dd7a68bcadf010f1c/CONTEXT.md), from my course-video-manager repo.
Which one is easier to read?
- BEFORE: “There’s a problem when a lesson inside a section of a course is made ‘real’ (i.e. given a spot in the file system)”
- AFTER: “There’s a problem with the materialization cascade”
This concision pays off session after session. This is built into /grill-with-docs. It’s a grilling session, but that helps you build a shared language with the AI, and document hard-to-explain decisions in ADRs. It’s hard to explain how powerful this is. It might be the single coolest technique in this repo. Try it, and see.
A shared language has many other benefits than reducing verbosity:
- Variables, functions and files are named consistently, using the shared language
- As a result, the codebase is easier to navigate for the agent
- The agent also spends fewer tokens on thinking, because it has access to a more concise language
#3: The Code Doesn’t Work
“Always take small, deliberate steps. The rate of feedback is your speed limit. Never take on a task that’s too big.”
David Thomas & Andrew Hunt, The Pragmatic Programmer
The Problem: Let’s say that you and the agent are aligned on what to build. What happens when the agent still produces crap? It’s time to look at your feedback loops. Without feedback on how the code it produces actually runs, the agent will be flying blind.
The Fix: You need the usual tranche of feedback loops: static types, browser access, and automated tests. For automated tests, a red-green-refactor loop is critical. This is where the agent writes a failing test first, then fixes the test. This helps give the agent a consistent level of feedback that results in far better code.
I’ve built a /tdd skill you can slot into any project. It encourages red-green-refactor and gives the agent plenty of guidance on what makes good and bad tests.
For debugging, I’ve also built a /diagnose skill that wraps best debugging practices into a simple loop.
#4: We Built A Ball Of Mud
“Invest in the design of the system every day.”
Kent Beck, Extreme Programming Explained
“The best modules are deep. They allow a lot of functionality to be accessed through a simple interface.”
John Ousterhout, A Philosophy Of Software Design
The Problem: Most apps built with agents are complex and hard to change. Because agents can radically speed up coding, they also accelerate software entropy. Codebases get more complex at an unprecedented rate.
The Fix for this is a radical new approach to AI-powered development: caring about the design of the code. This is built in to every layer of these skills:
/to-prdquizzes you about which modules you’re touching before creating a PRD/zoom-outtells the agent to explain code in the context of the whole system
And crucially, /improve-codebase-architecture helps you rescue a codebase that has become a ball of mud. I recommend running it on your codebase once every few days.
Summary
Software engineering fundamentals matter more than ever. These skills are my best effort at condensing these fundamentals into repeatable practices, to help you ship the best apps of your career. Enjoy.
Reference
Engineering Skills
I use daily for code work.
- diagnose — Disciplined diagnosis loop for hard bugs and performance regressions: reproduce → minimise → hypothesise → instrument → fix → regression-test.
- grill-with-docs — Grilling session that challenges your plan against the existing domain model, sharpens terminology, and updates
CONTEXT.mdand ADRs inline. - triage — Triage issues through a state machine of triage roles.
- improve-codebase-architecture — Find deepening opportunities in a codebase, informed by the domain language in
CONTEXT.mdand the decisions indocs/adr/. - setup-matt-pocock-skills — Scaffold the per-repo config (issue tracker, triage label vocabulary, domain doc layout) that the other engineering skills consume. Run once per repo before using
to-issues,to-prd,triage,diagnose,tdd,improve-codebase-architecture, orzoom-out. - tdd — Test-driven development with a red-green-refactor loop. Builds features or fixes bugs one vertical slice at a time.
- to-issues — Break any plan, spec, or PRD into independently-grabbable GitHub issues using vertical slices.
- to-prd — Turn the current conversation context into a PRD and submit it as a GitHub issue. No interview — just synthesizes what you’ve already discussed.
- zoom-out — Tell the agent to zoom out and give broader context or a higher-level perspective on an unfamiliar section of code.
- prototype — Build a throwaway prototype to flesh out a design — either a runnable terminal app for state/business-logic questions, or several radically different UI variations toggleable from one route.
Productivity
General workflow tools, not code-specific.
- caveman — Ultra-compressed communication mode. Cuts token usage ~75% by dropping filler while keeping full technical accuracy.
- grill-me — Get relentlessly interviewed about a plan or design until every branch of the decision tree is resolved.
- handoff — Compact the current conversation into a handoff document so another agent can continue the work.
- teach — Teach the user a new skill or concept over multiple sessions, using the current directory as a stateful teaching workspace.
- write-a-skill — Create new skills with proper structure, progressive disclosure, and bundled resources.
Misc
Tools I keep around but rarely use.
- git-guardrails-claude-code — Set up Claude Code hooks to block dangerous git commands (push, reset –hard, clean, etc.) before they execute.
- migrate-to-shoehorn — Migrate test files from
astype assertions to @total-typescript/shoehorn. - scaffold-exercises — Create exercise directory structures with sections, problems, solutions, and explainers.
- setup-pre-commit — Set up Husky pre-commit hooks with lint-staged, Prettier, type checking, and tests.
Similar Articles
@axichuhai: Whoa, a programming guru distilled his engineering experience into an open-source project that topped GitHub trending, with stars soaring past 90k+. The author is former Vercel engineer Mat, known for making complex tech easy to understand and involved in early Next.js development. He distilled his daily collaboration with Claud…
Introducing hello-agents, an open-source project that topped GitHub trending. It systematically organizes courses on AI and Agents from theory to practice, covering skills like Agentic RL, SFT, GRPO. Created by former Vercel engineer Mat, it distills engineering experience into 16 skills, such as Grim prompting techniques and red-green cycle testing methods.
@bkdgiffug: Still frustrated with AI writing code that makes things worse? Former Vercel engineer Matt Pocock has open-sourced his 18 battle-tested playbooks, specially designed to fix all kinds of AI collaboration failures. This guy who makes TypeScript crystal clear has turned his daily skill set into "brake pads" for AI programming — to let…
Former Vercel engineer Matt Pocock has open-sourced his AI coding agent skill set mattpocock/skills, containing 18 practical methods (such as /grill-me, /tdd, /caveman) to help developers better control AI-assisted coding and avoid writing bad code. It can be installed via npx skills or the Claude Code plugin.
@XAMTO_AI: An OpenAI founding member threw a CLAUDE.md onto GitHub, gaining 44k stars in a week, total stars hit 100k—even Stack Overflow didn't get that treatment. He observed that the biggest problem with AI programming is not that it can't work, but that it "loves to improvise": ask it to fix a ...
OpenAI founding member releases CLAUDE.md file, uses four principles to constrain AI coding assistant from over-improvising and complicating things, gaining 44k GitHub stars in a week.
@FakeMaidenMaker: Folks, we've dug up another AI engineering treasure: Hands-On-AI-Engineering. Just open-sourced, it already hit 2.3K stars. One repo stuffed with over 50 real AI projects that you can run directly. The most practical part is it doesn't talk about vague theories; each project is a complete little…
Recommends the freshly open-sourced GitHub repo Hands-On-AI-Engineering with 2.3K stars, containing over 50 runnable AI projects covering RAG, AI agent, OCR, and more. Each project provides full code and instructions, suitable for hands-on learning.
@yaohui12138: Karpathy released a GitHub open-source project that truly amazed me. The project is called andrej-karpathy-skills, with 130k+ stars on GitHub. I'd call it the most useful AI engineering project of 2026. The problem it solves is extremely precise: making Cl…
Karpathy released an open-source project called andrej-karpathy-skills, centered around a 4KB CLAUDE.md file containing 4 behavioral guidelines (Think Before Coding, Simplicity First, Surgical Changes, Goal-Driven Execution). It significantly reduces AI coding error rates (up to 90%), improving code quality and development efficiency.