Navigation menu

For developers who already work with Claude Code, Codex or Cursor

Stop reading every line your agent writes.

AI Developer Toolkit is the manual for working with Claude Code, Codex and Cursor. Its guides, Setup Pack and books show you how to hand over whole tasks and trust what comes back. You review the spec and the test report. The checks you set up review the diff.

Then $19.99 a month or $99.99 a year. Cancel from your account page. 30-day refund on a first purchase.

Guides for the three tools
400+
Recipes you can run today
100+
Book chapters, EPUB and PDF
328
MCP servers in the Setup Pack, checked 2026-10-02
31

English and Polish. Every page shows when it was last updated.

One task, start to finish, on a Level 4 setup.

Your attention goes to the spec before the build and the evidence after it. The diff goes to checks you set up once.

  • Written by you
  • Read by you
  • Agent and checks
  1. TicketIntent

    PAY-412

    Prorate plan changes made mid-cycle.

    You write

  2. SpecIntent

    spec.md

    Acceptance criteria, edge cases, files the agent must not touch.

    38 lines

    You write and read

  3. PlanIntent

    plan.md

    Seven steps. The agent proposes them, you approve before it builds.

    Agent writes

  4. BuildLoop

    +612 −148

    14 files changed in an isolated worktree, until the tests pass.

    Agent writes

  5. VerifyVerification

    • 212 tests passed
    • Stop hook passed
    • Review passed

    760 changed lines read

    Checks read

  6. Pull requestRelease

    #1184

    Opened with the test report attached. You read it against the spec.

    24 lines

    You read

Harness: permissions, sandbox, MCP servers and hooks, installed once
Lines you read62
Lines the checks read760
Drawing
One task on a Level 4 setup
Runs in
Claude Code, Codex, Cursor
Note
An illustration of the workflow, not a measurement

Engineering teams at these companies use it

  • Grupa Morizon-Gratka
  • Ringier Axel Springer
  • TasteRay

Creator of wondel.ai Skills — 2.3k+ stars on GitHub (opens in a new tab)

Your agent got faster. Your review queue didn’t.

Once the agent writes most of the code, you become its full-time reviewer. Every task ends in a diff you either read line by line or approve on trust.

“You are still responsible for your software, just as before.”

Andrej Karpathy, “Sequoia Ascent 2026 summary” (opens in a new tab), 2026-04-30
“Your life is diffs.”
Dan Shapiro on Level 3, the developer who reviews everything the agent writes. “The Five Levels”, 2026-01-23

Tue 09:12

You approve a 600-line diff you only skimmed, because reading it properly would take longer than writing it.

Tue 11:30

Your CLAUDE.md started as a paste from someone’s thread. Nobody on the team knows which lines still do anything.

Tue 14:05

Claude Code, Codex and Cursor each shipped changes this week. You heard about them from a post on X.

From our reader survey

“I switch between Claude Code, Cursor, VS Code, GitHub, and other agents, and often have to explain the same project or problem again.”

A reader, July 2026

What two years of telemetry from 22,000 developers shows

Faros AI, “AI Engineering Report 2026”, April 2026 (opens in a new tab)

+66%

epics completed per developer

+242%

incidents per pull request

+441%

median time a pull request spends in review

Output went up, and so did incidents and review time. The way out is checks that read the code for you, set up once and run on every task. That is what the toolkit teaches.

Know your level. Then climb one rung.

Dan Shapiro’s five levels describe how much of the work you hand to an agent, from spicy autocomplete to the dark factory. The guides are filed by level, so the next step is always named.

Not sure which row is yours?

The AI Dev Score asks 25 questions in about 8 minutes, names your level and the one move that takes you up a rung. Free.

Take the AI Dev Score
  1. 5

    The dark software factory

    “a black box that turns specs into software”

    WritesAI

    ReadsNobody

  2. 4

    The engineering team

    “write a spec” … “leave for 12 hours, and check to see if the tests pass”

    WritesAI

    ReadsTests, mostly

    Level 3 to 4 guides
  3. 3

    The developer

    “Your life is diffs.” … “almost everyone tops out here”

    WritesAI, most of it

    ReadsYou, full-time

    The valley, levels 2 and 3
  4. 2

    The junior developer

    “where 90% of ‘AI-native’ developers are living right now”

    WritesYou and the AI, paired

    ReadsYou, every line

  5. 1

    The coding intern

    “you offload specific, discrete tasks to your AI intern”

    WritesYou, mostly

    ReadsYou, all of it

  6. 0

    Spicy autocomplete

    “not a character hits the disk without your approval”

    WritesYou

    ReadsYou

Level names and quotations: Dan Shapiro, “The Five Levels” (opens in a new tab), danshapiro.com, 2026-01-23. The share figures are his. The ladder section of the toolkit is free to read: start with the ladder guide.

Start with the job in front of you.

Pick the situation you are in today. Each one has a place to start inside the toolkit.

  1. WhenThe agent’s diffs are bigger than your attention span.

    You get The Level 3 to 4 patterns: test gates, Stop hooks, sandboxes and review agents, worked out for Claude Code, Codex and Cursor, so a task ends with evidence instead of a wall of code.

    Verification guides

  2. WhenYour agent config is a paste from someone else’s repo.

    You get The Setup Pack. Answer 9 questions about your stack and download AGENTS.md, CLAUDE.md, Cursor rules and MCP config from 31 checked servers, plus an install prompt your agent runs to add the Stop hook and check every setting against what you have installed.

    The Setup Pack

  3. WhenThe task is too big to hand over in one prompt.

    You get The brief chain: intent.md, spec.md, plan.md. The agent works from a spec you wrote and approved, in the same order in Claude Code, Codex and Cursor.

    Spec-first workflow

  4. WhenThree tools ship changes every week and you can’t keep up.

    You get Guides that cover Claude Code, Codex and Cursor side by side, a last-updated date on every page, and What’s New, a feed of what changed and which guides changed with it. What’s New is free for everyone.

    What’s New

  5. WhenYour team works at five different levels.

    You get Shared brief templates, repository context and verification gates for every seat, a reading path for tech leads, and the free Tech Lead Scorecard to place the whole team on the ladder.

    The Team plan

  6. WhenThe agent answered with a flag that doesn’t exist.

    You get Codex has no --approval-mode flag. The real one is --ask-for-approval, with the values on-request and never (Codex CLI 0.159.2, checked 2026-10-02). The command reference lists stale flags the same way, each with the release it was checked against.

    The command reference

Everything in the Solo Developer plan.

One subscription covers the library, the Setup Pack and the books. The ledger prices the pack and the library titles.

Setup Pack, sold alone$49

The four-title book library, at a comparable ebook’s price$191

Priced items$240

A year of Solo Developer$99.99

Books: each library title is priced at $47.99, explained in the books section below. The guide library and the three role paths are not counted.

  • The guide library

    400+ guides and 100+ recipes for Claude Code, Codex and Cursor, filed by level, in English and Polish. Every page shows when it was last updated.

    Includedin the plan

  • The Setup Pack

    Nine questions about your stack, about 20 minutes, one zip generated for your repository.

    • AGENTS.md
    • CLAUDE.md
    • .cursor/rules
    • .mcp.json
    • TESTING.md
    • INSTALL-PROMPT.md

    $49on its own

  • Seven books

    The four-title library, 328 chapters on 3,057 pages, plus three reading paths for developers, tech leads and CTOs. EPUB and PDF, English and Polish. Downloads are yours to keep. See the shelf

    $191four library titles

  • Free for everyone, no plan needed: the Start here and ladder guides, the three scorecards, the MCP server that lets your agent search the library, What’s New, the Discord for questions, and the public GitHub repo where you report a mistake.

Seven books, included in every plan.

The four library titles run to 3,057 pages. At $47.99 each, the typical price of a comparable ebook, they would cost $191. The three role paths reuse those pages in the order your role needs, so they add nothing to that total.

Cover of The AI Engineering Handbook

Library1 of 4

The AI Engineering Handbook

Cursor, Claude Code & Codex — The Craft of Shipping with AI

Read it once and the loop stops depending on which tool is open: you write the spec, let a failing test set the stop condition, and read the evidence at the gate instead of every diff.

Chapters
117
Pages
1,224
Formats
EPUB, PDF
Editions
EN, PL

Typical price of a comparable ebook

$47.99

Included in every plan

Cover of AI Engineering with Cursor

Library2 of 4

AI Engineering with Cursor

A Practical Guide to AI-Native Development in the Cursor IDE

Pick the smallest Cursor surface that can finish each job, from Tab to Cloud Agents, and trust the result because a test, CI and Bugbot checked it, not because the diff looked right.

Chapters
72
Pages
611
Formats
EPUB, PDF
Editions
EN, PL

Typical price of a comparable ebook

$47.99

Included in every plan

Cover of AI Engineering with Claude Code

Library3 of 4

AI Engineering with Claude Code

A Practical Guide to Agentic Development in the Terminal

What it takes to leave a terminal agent running without you: a CLAUDE.md that names the gates, a permission mode for each kind of session, and hooks that enforce what a prompt only asks for.

Chapters
81
Pages
741
Formats
EPUB, PDF
Editions
EN, PL

Typical price of a comparable ebook

$47.99

Included in every plan

Cover of AI Engineering with Codex

Library4 of 4

AI Engineering with Codex

Building with OpenAI Codex from the CLI to the Cloud

Which Codex surface fits the job, from the CLI and desktop app to the IDE extension and the cloud, with a sandbox and approval policy you chose on purpose.

Chapters
58
Pages
481
Formats
EPUB, PDF
Editions
EN, PL

Typical price of a comparable ebook

$47.99

Included in every plan

Cover of Agentic Engineering for Developers

Role path1 of 3

Agentic Engineering for Developers

A reading path up the autonomy ladder: from your first merged agent change to reviewing evidence instead of diffs

Write the instructions file, set permissions, give the agent a feedback loop it runs itself, then approve changes on their evidence instead of reading every line.

Steps
16
Further reading
2
Pages
199
Editions
EN, PL

Not counted in the value

Built from the library’s pages, in the order a developer needs them.

Included in every plan

Cover of Agentic Engineering for Tech Leads

Role path2 of 3

Agentic Engineering for Tech Leads

Move the whole team up one rung: harness, verification design and review capacity

Place the team on the ladder, set a baseline, standardize the harness, design verification and run code review as a queue with limits.

Steps
16
Further reading
4
Pages
245
Editions
EN, PL

Not counted in the value

Built from the library’s pages, in the order a tech lead needs them.

Included in every plan

Cover of Agentic Engineering for CTOs and Executives

Role path3 of 3

Agentic Engineering for CTOs and Executives

The brief, the decisions you own and the operating model for agent-built software

The dated evidence, the cost of an accepted change, a pilot with a control group, autonomy by risk class, agent security, spend and the board update.

Steps
19
Further reading
8
Pages
304
Editions
EN, PL

Not counted in the value

Built from the library’s pages, in the order a CTO or executive needs them.

Included in every plan

1 / 7

Typical price: the median list price of ten ebooks on AI-assisted and agentic coding from Leanpub, Manning, O'Reilly, Packt and Pragmatic Bookshelf, checked 2026-10-10 (range $24.95 to $59.99). Page counts are the English A4 PDF editions.

Your first week, planned.

  1. Day 1

    Generate your Setup Pack.

    Answer nine questions and download the zip. Your agent runs INSTALL-PROMPT.md and checks each setting against the version you have installed. About 20 minutes.

  2. Days 2 to 5

    Hand over one real ticket.

    Write intent.md and spec.md for something you were going to build anyway. Let the agent plan, build and stop on the tests. Read the report, then the code if you want to.

  3. Day 7

    Decide with evidence.

    You have one task done the new way. Keep going at $19.99 a month or $99.99 a year, or cancel from your account page before the end of day 7.

Start 7-day free trial Unlock every guide

Then $19.99 a month or $99.99 a year. Cancel from your account page. 30-day refund on a first purchase.

Who checks this, and how.

Every guide written and maintained by Michał Jaskólski.

  • 26 years building web apps
  • 2 IPOs
  • AI Lecturer at SWPS University
  • Using Cursor since 2023, Claude Code and Codex since early beta
  • CTO at TasteRay (built with AI)
codex, checked 2026-10-02
# what a model suggested
codex --approval-mode never
# what Codex CLI 0.159.2 accepts
codex --ask-for-approval never
  • Every page carries a date.

    Each guide shows when it was last updated, so you always know how old what you are reading is.

  • The command reference names the releases it was checked on.

    It states the Claude Code and Codex versions behind it, and its table of stale flags gives the release each row was checked against.

  • The Setup Pack catalogue is verified against its sources.

    The 31 MCP servers in its catalogue were last verified against their upstream sources on 2026-10-02, 31 of 31. A weekly CI run is set up to repeat the check.

  • Numbers come from third parties, with a date.

    Every statistic on this page names its publisher and when it was published. No multipliers that nobody measured.

  • Mistakes are reported in public.

    Every guide has an Improve this page button that opens an issue in the public GitHub repo. When the author confirms it, the guide is corrected and its date changes.

One price for you. One for your whole team.

Solo and Team start with seven free days. Prices in USD; VAT is added where it applies.

Solo Developer

Start here

For a developer who wants to stop reading every line the agent writes.

$19.99$99.99 per monthper year

or $99.99 a year, 58% less than monthly

$8.33 a month, 58% less than $19.99 monthly

Start 7-day free trial Manage your plan

7 days free, then the plan you picked

  • The Setup Pack, generated for your stack ($49 on its own)
  • 400+ guides and 100+ recipes, each with a last-updated date
  • All seven books, EPUB and PDF, English and Polish
  • The Level 3 to 4 patterns: sandboxes, Stop hooks, verification gates

Team

One setup for the whole team, from the first day.

$199.99$499.99 per monthper year

or $499.99 a year, 79% less than monthly

79% less than twelve payments of $199.99

Start team trial Manage your plan

7 days free, one flat price, no per-seat billing

  • Everything in Solo Developer, for every engineer
  • Unlimited seats, by email domain or by named address
  • Shared brief templates, repository context and verification gates
  • One company invoice with VAT
  • Direct email support from the author

Enterprise

An audit of where your teams really are, a measured pilot, and someone accountable for adoption.

Custom

Priced by organization size and pilot scope

Book an intro call (opens in a new tab)

Delivered personally by the author

  • Everything in Team, unlimited seats
  • An audit of how each team works with agents, from plan to maintenance
  • The brief chain and CI evidence installed in your repositories
  • Workshops and a pilot team, adoption measured from day one
  • 7 days free

    Try Solo or Team before the first charge.

  • 30-day refund

    On a first purchase, for any reason. Email support and it is refunded.

  • Cancel from your account

    Access runs to the end of the period you paid for. Downloaded books stay yours.

  • VAT invoice for every charge

    Polar, the merchant of record, sends it right after payment.

Solo, billed yearly$8.33 / month

1 hour at $80k a year≈ $38

Break-even13 minutes a month

Save 13 minutes a month and the plan has paid for itself.

$8.33 divided by $38 an hour comes to 13 minutes. With your own hourly rate the number moves.

Who it is for.

A good fit if

  • You use Claude Code, Codex or Cursor at work every week.
  • You want to hand over bigger tasks and still sign off on what ships.
  • You lead a team and want everyone working at the same level.

Probably not a fit if

  • You want a promise that you’ll ship several times faster. The studies we cite show output and incidents rising together, so we make no such promise.
  • You haven’t tried a coding agent yet. Start with the free ladder guide.
  • Your team wants a plugin and won’t change how it works.

Questions people ask before starting.

Something else? Write to support@developertoolkit.ai and the author answers.

Can’t I just ask the AI how to set this up?

You can, and the answer will often sound right. Models invent flags: ask for Codex approval settings and you may get --approval-mode, which does not exist. The toolkit keeps a command reference that names the release each stale flag was checked against, and every guide shows when it was last updated.

How is this different from the official docs?

Official docs describe one tool’s features. The toolkit describes the workflow across all three, filed by level: what to set up, in what order, and how to check the result.

I only use one of the three tools.

That works. Each guide says which tool it covers, and the workflow is the same order in all three, so moving to another tool later costs you little.

What happens when the trial ends?

The plan you picked starts: $19.99 a month or $99.99 a year for Solo Developer. Cancel from your account page before the end of day 7 and you are not charged.

Do you offer refunds?

Yes, 30 days from your first purchase, for any reason. Email support@developertoolkit.ai and it is refunded. The guarantee applies to first-time purchases, as set out in the Terms.

Do I get an invoice for my company?

Yes. Polar, the merchant of record, issues an invoice for every charge with the VAT or sales tax for your country. It arrives right after payment.

Will agents replace me?

Your work moves to the spec, the architecture and the stop conditions. Karpathy says you are “in charge of taste, engineering, design, and whether the system makes sense.” The toolkit is about doing those well.

Hand your next ticket to a Level 4 setup.

“level 2, and every level after it, feels like you are done. But you are not done”

Dan Shapiro, “The Five Levels” (opens in a new tab), 2026-01-23

Then $19.99 a month or $99.99 a year. Cancel from your account page. 30-day refund on a first purchase.