explainx.ai0k
TrendingAI News TodayPathwaysSkills
Pricing
explainx.ai

Upskill in AI — 16 free pathways, live workshops & bootcamps, and 50+ courses from practitioners. Plus the skills, tools, and MCP servers to practice on.

follow us

follow on google

Add explainx.ai as a preferred source

corporate training

support@explainx.ai

get started

Find your pathTake Free Evaluation

community

Join the community

learn

mind: share how you thinkpathways — start freeworkshopsbootcampscoursescompare Explainxcertificationsmock testsexplainx universitycorporate traininglearn skills & mcp

discover

skillsmcp serversexplainx mcptoolsmdx readeragentsllmsdesignsdictionarypeopleagi trackerfelony benchranks

company

aboutvisionmissionteaminstructorsteach on explainxpartnershipscommunityhackathonscareers

content

daily AI newsstate of AI — live resultsblogreleasespromptsgeneratorsresource libraryfor LLMsexplainx.ai kids

solutions

all solutionsdeveloper upskillingmarketing upskillingproduct manager upskillingleadership upskilling

newsletter · weekly

Get AI news, tools, and insights in your inbox.

supportcontactprivacytermsdata rightshow we create contentsubmission guidelines

© 2026 AISOLO Technologies Pvt Ltd

explainx.ai

On this page

  • TL;DR: the questions people are asking
  • Why effort per subagent matters
  • How to use it
  • What we could not verify
  • The cost picture
  • Risks and failure modes
  • When not to bother
  • A worked example: auditing payment calls
  • What this means for what you build or pay
  • Related reading
← Back to blog

explainx / blog

Claude Code Subagents Now Take an Effort Level: How to Run Cheap Scouts and Careful Reviewers

Claude Code, Subagents, Effort Levels, Anthropic, Guides

Claude Code v2.1.292+ lets you ask for subagents at a set effort level. How to pair low-effort scouts with a high-effort reviewer, with prompts and cost tips.

Oct 7, 2026·9 min read·Yash Thakker
add explainx.ai
go deep
Claude Code Subagents Now Take an Effort Level: How to Run Cheap Scouts and Careful Reviewers

A one-line announcement landed on October 7, 2026 that changes how you can budget a multi-agent run. Lydia Hallie, who works on Claude Code at Anthropic, wrote: "You can now ask Claude to run subagents at a specific effort level! Make sure you're on v2.1.292+." The attached screenshot shows the prompt style:

Find every payments API call with low effort subagents, then have a high effort one check the error handling.

That example is the whole idea in one sentence: cheap, fast subagents do the wide search, and one careful subagent does the part that needs judgment. This guide explains why that matters, how to use it, what we could not verify from the announcement, and how to keep costs under control. We worked from the announcement and Anthropic's existing effort and subagent documentation as summarized in our earlier posts, not from release notes for 2.1.292, which we did not have.

Weekly digest3.5k readers

Catch up on AI

Curated AI updates on agents, skills, and MCP — delivered to your inbox. Unsubscribe anytime.

TL;DR: the questions people are asking

table · 2 cols
QuestionShort answer
What is new?You can ask for subagents at a chosen effort level.
Which version?v2.1.292 or later.
How do you ask?In plain language, per the announcement's example.
Why do it?Spend reasoning where it matters, save it where it does not.
Does it save money?It can trim cost on scouting steps. Subagents still add overhead.
What is unverified?Exact syntax, which levels subagents accept, and defaults.

Why effort per subagent matters

Effort is Claude Code's control for how much reasoning the model spends on a turn. Our commands reference lists /effort [low|medium|high|xhigh|max|ultracode|auto], with max and ultracode limited to the session. Until now, effort mostly applied to the session as a whole. Setting it high made everything slower and more expensive, setting it low risked sloppy work on the hard parts.

Multi-agent work makes that trade-off sharper, because a run is a mix of very different jobs.

table · 3 cols
JobReasoning neededBetter effort
Find every call site of an APILow: pattern matching and readingLow
Summarize what each file doesLow to mediumLow or medium
List TODOs and dead codeLowLow
Check error handling for correctnessHigh: edge cases, failure modesHigh
Review a concurrency changeHighHigh
Judge a security-sensitive diffHighHigh or higher
Decide an architecture trade-offHighHigh, with a human in the loop

Per-subagent effort lets one request express that shape: wide and cheap on the left, narrow and careful on the right. It is the multi-agent version of the advice in our post on model versus effort: the model sets what Claude knows, the effort sets how hard it tries.

How to use it

Start from the announcement's pattern and adapt it. The structure is always the same: name the work, name the effort, and say how the results combine.

1. Search, then review.

text
Find every place we call the payments API using low effort subagents,
one per top-level package. Then run a high effort subagent over the
combined list to check that each call handles timeouts, retries and
idempotency keys. Report only the calls that are unsafe.

2. Parallel summaries, one synthesis.

text
Use low effort subagents to summarize each service directory in two
sentences. Then have one high effort subagent read the summaries and
propose where the module boundaries are wrong.

3. Cheap triage before an expensive fix.

text
Use low effort subagents to reproduce each failing test and classify
the cause as flaky, environment or real. For the real failures only,
use a high effort subagent per failure to propose a fix.

4. Review a large diff in slices.

text
Split this diff by directory. Use medium effort subagents to review each
slice for style and obvious bugs. Use one high effort subagent to look
for cross-slice problems such as changed interfaces.

A few habits improve results. State the effort once per phase, not per sentence, so the instruction is unambiguous. Keep the final, careful step to a single subagent unless the work truly splits, because high-effort parallel agents are where cost climbs. And say what you want back: a short list, a table, or a verdict, so cheap subagents do not return pages of raw output that the expensive one must then read.

What we could not verify

The announcement is brief, and some practical questions are open.

  • Syntax. The example uses plain language. We do not know whether there are also settings, flags or a frontmatter field in subagent definitions for effort. If you define custom subagents, check the format in your version's documentation. Our guide to Claude Code subagents and multi-agent workflows describes the definition format as of its writing.
  • Which levels apply. The session accepts low, medium, high, xhigh, max, ultracode and auto. Whether subagents accept all of them, or a subset, is not stated.
  • Defaults. What effort a subagent uses when you say nothing, and whether it inherits the session level, we did not confirm.
  • Interaction with models. Whether a subagent can use a different model as well as a different effort in the same request is outside what the post says.
  • Reporting. Whether the session shows each subagent's effort and cost separately is unknown.

The reliable way to find out is to test. Run claude --version, update if needed, and try a small, harmless prompt with two effort levels, then compare the transcript and /usage.

The cost picture

Subagents are not free. Our measurements in Do subagents actually use more usage? showed that each subagent builds its own context, so total tokens rise with parallelism even when the wall-clock time falls. Per-subagent effort gives you a lever on part of that bill.

A rough way to reason about it:

  • Scouts dominate in count. A search across 40 files might spawn many subagents. Making those low effort reduces the reasoning tokens in the part of the run that has the most agents.
  • The reviewer dominates in depth. One high-effort subagent costs more per turn, but there is only one, and it works on a compact input.
  • The synthesis step is where waste hides. If cheap scouts return verbose output, the expensive reviewer pays to read it. Ask scouts for terse, structured results.

For a sense of what a single task costs on the current top model, see What a Claude Code task costs on Opus 5.5, and for the broader question of when more effort is worth it, our post on effort levels and plan mode.

Risks and failure modes

Under-powered scouts miss things. A low-effort subagent can skim and skip a call site hidden behind a wrapper. Fix: make the scouting step produce evidence (file and line), and have the reviewer spot-check coverage with a different method, such as a grep.

A confident wrong summary. If the high-effort step trusts a low-effort summary, an error propagates. Fix: have the reviewer read the underlying code for anything it flags, not only the summaries.

False economy. Saving tokens on a step that later needs a redo costs more. Fix: use medium for steps that are not clearly mechanical.

Permission sprawl. More subagents means more tool calls. Keep your permission mode appropriate, as in our guide to Claude Code permission modes, and avoid blanket approvals for parallel runs.

Inconsistent standards. Different effort levels can produce differently styled output. Specify the format.

When not to bother

Skip the pattern for small tasks. If a change touches three files, a single agent at medium effort is simpler and often cheaper than orchestrating subagents. Use the multi-agent, mixed-effort approach when the work is wide (many files or services), the steps are separable, and there is a clear difference between a mechanical phase and a judgment phase.

A worked example: auditing payment calls

Consider the announcement's own scenario, a codebase with a payments API called from a dozen services. A sensible run has three phases, and each gets its own effort.

Phase one, discovery (low effort). One subagent per service lists every call site with file, line and the function name, plus whether the call is wrapped in a retry. The output is a flat table, nothing more. Because the work is pattern matching, low effort is enough, and running a dozen of them in parallel keeps wall-clock time short.

Phase two, consolidation (the main agent). The table is merged and deduplicated. This is cheap, and it is the natural place to sanity-check coverage, for example by comparing against a plain text search for the client library name to catch anything the scouts missed.

Phase three, review (high effort). One subagent reads the actual code around each call that lacks a retry or timeout, and answers specific questions: what happens on a network error, a partial response, a duplicate submission, an expired token? It returns a short list of unsafe calls with a suggested fix for each.

The result is a run where most of the tokens go to reading code, and the expensive reasoning is concentrated on a few dozen lines that matter. Compared with running everything at high effort, you spend less on the search. Compared with running everything at low effort, you avoid the risk of a shallow review on the part where mistakes are costly.

After the run, check three things: whether the scouts' call-site list matches a simple search, whether the reviewer's flagged items are real when you read them, and what the session cost in /usage. Those numbers tell you whether the split paid off on your codebase and whether to keep it as a standing habit.

What this means for what you build or pay

For individual developers on limited plans, the feature is mostly about allowance: you can keep broad searches cheap and spend your high-effort budget on the review that decides whether the change is safe. For teams, it suggests a convention worth writing into your project instructions: scouts at low, reviewers at high, final sign-off by a human. For people building their own agent harnesses, it confirms a design direction that other tools are taking: effort and model as per-task knobs, not session-wide settings. If you track how agent tooling is evolving, it fits the broader move we described in the terminal era and the agent as the new primitive.

Related reading

  • Claude Code commands: complete reference
  • Claude Code subagents and multi-agent workflows
  • Do subagents actually use more usage?
  • Model versus effort: knowing more versus trying harder
  • What a Claude Code task costs on Opus 5.5
  • Effort levels and killing plan mode
  • Claude Code permission modes explained
  • Hermes Agent manual subagent control

Primary: Lydia Hallie's announcement on X (October 7, 2026), including the example prompt in the attached screenshot · Claude Code release notes for v2.1.292, which should be checked for exact behavior

Details are accurate as of October 7, 2026 and come from the announcement and our earlier documentation summaries. We did not read the v2.1.292 release notes, so syntax, supported levels and defaults are unverified. Test on a small task before relying on it.

Spotted something out of date? Let us know.
Yash Thakker

Written by

Yash Thakker

Yash is an AI expert with over 300K learners. Join his workshops →

View Yash Thakker in People in AI →

Related posts

Aug 23, 2026

Claude Code Effort Showing 10/100? It Was a Display Bug, Not a Downgrade

Claude Code users on Hacker News and X noticed the numeric effort value next to their session drop to 10 out of 100 — the number "low" used to show — while still selecting "high." Anthropic's Thariq confirmed it was a serving-config experiment that remapped the display scale, not a change to how much work Claude actually does. explainx.ai breaks down the thread, the fix, and how to verify your own sessions.

Aug 15, 2026

Why Does Claude Opus 5 Feel Worse to Work With? The HN Debate

"Why does Opus 5 feel worse to work with?" hit 778 points and 717 comments on Hacker News this week. The original post's theory: reinforcement learning from verifiable rewards trains models to commit to an answer instead of pausing to ask, and that trade-off shows up as a model that makes bold assumptions instead of checking them. explainx.ai breaks down the thesis, the recurring complaints from the thread, and how to prompt around it in Claude Code.

Aug 7, 2026

Why Developers Say Claude Opus 5 Over-Engineers Simple Tasks

A widely-upvoted r/ClaudeAI thread from August 6, 2026 crystallized a complaint builders had been trading for weeks — Claude Opus 5 writing its own elaborate briefs, then executing far past the original ask. explainx.ai breaks down the specific complaints and the six workaround patterns practitioners are actually using.