The frame
Vibe coding, vibe engineering, and Human-Driven Development; the four debts; and the production-debugging test every merge must pass.
Human-Driven Development with Coding Agents
A working method for using coding agents without giving up the engineering decisions that keep a product maintainable.
Full day · Half-day version available
Engineers and technical founders using Claude Code, Codex, pi, or similar coding agents in real projects.
Bring a laptop with Claude Code, Codex, or pi installed and authenticated.
Projects rot in the gap between generation speed and review capacity: cognitive debt, verification debt, architectural debt, and the loss of judgment that follows when the tool does the work without teaching the person directing it.
This workshop turns Human-Driven Development into repeatable practice. We work hands-on with Claude Code, Codex, or pi, while keeping the durable controls in the repository: instructions, automated checks, review gates, work ledgers, and release evidence.
The team structures a repository so any coding agent inherits its standards, moves enforceable rules from prose into machinery, runs implementation and review as separate turns, and builds verification around the work before it reaches users.
The goal is not to type more. It is to decide what the agent may write before it writes it, and to preserve enough context and evidence that a person can still stand behind the result.
Vibe coding, vibe engineering, and Human-Driven Development; the four debts; and the production-debugging test every merge must pass.
Expand the deterministic surface, then use canonical documentation, repository instructions, glossaries, memory, and reusable workflows to carry guidance machines cannot enforce. Map the shared practice onto Claude Code, Codex, and pi through a compact harness translation guide.
One active queue, explicit implement-and-review alternation, approval gates for consequential changes, and a record of where execution departed from plan.
Separate orchestration from implementation, choose the cheapest adequate model for each task, compare how Claude Code, Codex, and pi delegate work, and trust recorded evidence over self-reported results.
Contract tests on fixtures, behaviour evaluations on representative cases, and invariant checks on the shipped deliverable itself.
Blind architecture review, dated findings, materiality-based human attention, deterministic release commands, and Git discipline across concurrent agent sessions.
Using Claude Code, Codex, or pi, the team receives a deliberately vibe-coded feature with representative defects, wraps it in the controls module by module, and watches those controls expose what is wrong.
Tell us what your team is working with and what should be different afterwards.