Skip to content

Tech lead track: move the whole team up one rung

The tech lead track is the reading path for tech leads who need a whole team, not one power user, to ship with Claude Code, Codex, or Cursor: one engineer merges five agent pull requests a day, two refuse the tools, and review is the slowest part of delivery. It covers three jobs: standardize the harness, own verification, and run review capacity.

What is a tech lead’s job when agents write the code?

Section titled “What is a tech lead’s job when agents write the code?”

Your job moves from writing code to designing the system that decides whether code ships:

Standardize the harness. The harness is what the team shares around the model: the instructions file, permissions, MCP servers, skills, and one command that runs tests, lint, and types. Keep it in the repository; review it like code.

Own verification design. Cheap generation moves the bottleneck to checking. DORA 2025 (Google Cloud, 23 September 2025) warns that without controls like strong automated testing, “an increase in change volume leads to instability.” You choose the tests, fitness functions, and acceptance criteria that prove a change correct.

Run review capacity. Review is a queue with fixed capacity: Faros AI’s AI Engineering Report 2026 (April 2026; 22,000 developers, vendor telemetry) measured throughput up 33.7%, median time in review up 441.5%, and incidents per pull request up 242.7%. You decide which changes a person reads and which merge on evidence.

Baseline first, transfer trust last: evidence-based review without strong tests only moves the risk.

Each step is marked Free or Subscribers; the three ladder steps are free. Read them in order, then skip any step whose exit evidence your team already has.

  1. Why now: One frame for maturity, process and capability before you change how eight people work.

    Done when: Your team placed on the map, loop by loop.

  2. Why now: Most teams stall at Level 3; know its review ceiling before you plan past it.

    Done when: The team knows how many diffs a week it can really review.

  3. Why now: The standard the team moves to: evidence first, code only for escalation classes.

    Done when: Escalation classes agreed with the team.

  4. Why now: Set the baseline before the rollout, or you will not be able to show what moved.

    Done when: Four weeks of baseline data for the metrics you will report.

  5. Why now: Senior refusal quietly caps the whole team; bring the skeptics in before the rollout.

    Done when: Each senior has shipped one real change with an agent, paired with you.

  6. Why now: Roll out one repository at a time, as transitions on the ladder, not a big bang.

    Done when: A roadmap with one repository, one loop and one exit criterion per step.

  7. Why now: Agents amplify what the repository already is; make it legible and testable first.

    Done when: A one-command bootstrap, a fast test suite and a current AGENTS.md.

  8. Why now: One shared, versioned rules core instead of eight private setups.

    Done when: Shared rules in the repository, reviewed like code.

  9. Why now: Agents do well on well-shaped work; shape the backlog so they get it.

    Done when: A ticket template with acceptance criteria and a risk class.

  10. Why now: Trusting tests you did not read needs a measure of how strong they are.

    Done when: A mutation or oracle-strength score on the repositories agents touch.

  11. Why now: Define what every agent pull request must prove, in one manifest.

    Done when: The evidence bundle required by CI on agent pull requests.

  12. Why now: Maintainability is checked by fitness functions, not by reading the code.

    Done when: Two architecture rules enforced in CI.

  13. Why now: Review by evidence and risk class, so review capacity stops being the bottleneck.

    Done when: Review time per agent pull request down, escaped defects flat.

  14. Why now: Run the review queue as a system when agents open most of the pull requests.

    Done when: A work-in-progress limit and a queue-age alert on the review queue.

  15. Why now: Move the team from reading every diff to trusting the evidence, deliberately.

    Done when: One risk class merged on evidence alone for a month without an incident.

  16. Why now: Juniors still need to grow judgment when agents write the code.

    Done when: A growth plan for each junior with review and design work built in.

Which scorecard question does each step move?

Section titled “Which scorecard question does each step move?”

Use this as a four-week plan: pick one Harness, one Verification, and one Review row, then re-take the scorecard. Numbers follow the answer key.

StepQuestion it movesExit evidence
Frame: One mapYour bandTeam placed on the ladder
Review: Level 3, you review the diffsQ6, review standardsWeekly review capacity known
Verification: Evidence instead of codeQ8, plausible-but-wrong codeEscalation classes agreed
Baseline: Metrics frameworksQ11, measuring impactFour weeks of baseline data
People: Skeptics and seniorsQ12, skeptics and seniorsEach senior shipped one agent change
Harness: Adoption roadmapQ10, how adoption happenedOne pilot repository, one exit criterion
Harness: Agent-ready codebaseQ2 and Q4, context and onboardingOne-command bootstrap, current instructions
Harness: Shared agent rulesQ1, a common agent configRules in the repository, reviewed like code
Harness: Agent-ready backlogQ5, consistent qualityTicket template with a risk class
Verification: Oracle strengthQ8, plausible-but-wrong codeMutation score on agent-touched code
Verification: Evidence bundleQ7, automated gatesCI requires the bundle
Verification: Fitness functionsQ5, consistent qualityTwo architecture rules enforced in CI
Review: Agent PR reviewQ6 and Q15, standards and PR loopReview time down, escaped defects flat
Review: Review queueQ15, the PR and merge loopWork-in-progress limit and queue-age alert
Review: Trust transferQ24, what AI cannot touchOne risk class merged on evidence
People: Junior developersQ18, levelling up the teamA growth plan per junior

Where the shared harness lives in each tool

Section titled “Where the shared harness lives in each tool”

In every tool, the team’s setup belongs in the repository, not on laptops.

  • Instructions: CLAUDE.md in the repository. Claude Code falls back to AGENTS.md from v2.1.277 on the latest channel, not yet on stable.
  • Permissions: the shared .claude/settings.json, which ignores a defaultMode of auto or bypassPermissions.
  • MCP servers: claude mcp add --scope project, checked in.
  • Review: /code-review locally, anthropics/claude-code-action@v1 in CI. Checked against Claude Code 2.1.283 on 26 September 2026.
  • One power user carries the numbers. The team average can hide engineers still at Level 1. Recovery: pair each holdout with the power user on a real ticket.
  • Review time rises after the rollout. Recovery: cap agent pull requests in progress and route changes by risk class before adding reviewers.
  • The harness drifts. Recovery: change instructions and settings only through pull requests; re-run the audit prompt monthly.
  • Trust transfers before the oracle is strong. Recovery: return that risk class to line-by-line review until oracle strength holds.

Frequently asked questions

What is the tech lead track?

An ordered reading path for tech leads and heads of team who are moving a whole team, not one power user, up the autonomy ladder. It opens with the free Tech Lead Scorecard as a baseline, then runs through three jobs: standardizing the harness, owning verification design, and running review capacity. Every step names the scorecard question it moves and the evidence that says the team can move on.

Why does the track start with the scorecard?

The scorecard takes about 8 minutes and scores the team on 25 questions. Its answers show which steps the team already meets, and re-taking it after four weeks shows which section scores moved. The answer key maps every question to the page that moves it.

Which steps are free?

The three ladder steps at the start of the path are free. Every step is marked Free or Subscribers.