Put Claude's Juniors to Work
Subagents: one line for your bill, one line for your chat — Claude adds both for you
Start Here: Paste This
Here from the juniors video? This is the line. Paste the prompt into Claude Code exactly as written — nothing to fill in. Claude adds the staffing rule to this project's CLAUDE.md, shows you what it wrote, and names the first job it would now delegate. From then on, small mechanical work goes to juniors on smaller, cheaper models and the big brain sticks to the thinking.
Never heard of CLAUDE.md? It's the document Claude reads before every job — the prompt above creates it if you don't have one, and the full guide to it is here.
What's Actually Happening
By default it's one Claude doing every part of a job personally — the thinking and the grunt work, all on the biggest model you've got, all piling into one memory. That's the senior partner renaming files, and it's why long chats strain and limits vanish.
But Claude can manage helpers — Anthropic calls them subagents. It hands a piece of work to a junior copy running in its own separate memory, on a smaller model where the job allows, and gets back just the result. The partner decides, a junior does it, the partner checks it after — you can simply ask for one in plain language, or use the lines on this page to make it standing behaviour.
You'll know it's working when a little task spins off to the side instead of Claude opening file after file in your chat.
The Two Lines
The staffing line — for your bill
Small mechanical jobs go to juniors on cheaper models, reviewed before they land. Tokens are the new billable hours — as AI takes on more of your work, the gap between one expensive brain doing everything and a staffed team is real money. It's the prompt at the top of this page.
The reading line — for your chat
Heavy reading — piles of files, logs, research — happens in a junior's memory instead of your conversation, so your chat stays lean and lasts far longer before it strains (that strain is the context window filling up). This prompt adds it:
Using The Claude App, Not Claude Code?
The feature lives in Claude Code, but the idea works anywhere: you play the junior's memory yourself. When a big side-quest comes up, do it in a separate chat and paste just the conclusion back into your main one — the heavy reading lives and dies in the side chat, and your main conversation only ever carries the answer.
The Honest Bit
Juniors still use tokens — the reading isn't free, and delegating a quick one-file edit is just overhead. The staffing line saves the premium you stop paying for the big model's time on jobs that never needed it; the reading line saves your chat from re-reading a pile of files on every reply. And CLAUDE.md is a strong nudge, not an iron rule — specific instructions like these stick far better than a vague "be efficient".