
Why Coding Agents Do Not Eliminate Expertise
A review of Anthropic's research: delegation, verification, and recovery when using Claude Code

A review of Anthropic's research: delegation, verification, and recovery when using Claude Code
A review of Anthropic's research: delegation, verification, and recovery when using Claude Code
Original redraw of Anthropic's research frame
Redrawn from Appendix · Sample · p. 2
Original redraw of the measurement pipeline
Redrawn from Figure 1 and Appendix · Work Mode
Redrawn from Figure 1
Redrawn from Figure 2
Redrawn from Who decides what
Redrawn from Level of expertise
Redrawn from Table 1 and Appendix · User expertise
Original diagram of possible construct overlap
Redrawn from Figure 3
Redrawn from Figure 4
Redrawn from Figure 4 and Appendix · Task value
Redrawn from Table 2 and Appendix · Derived measures
Redrawn from Figure 5
Redrawn from Figure 5 · troubled sessions
Redrawn from Figure 6
Redrawn from Appendix · Validation and robustness
Supported
Expertise tracks outcomes
Experts delegate more
Gradient survives controls
Not established
Training causes the same effect
More actions mean better work
Result generalizes to all agents
Original model: framing × verification × recovery
Build in people
Precise task framing
Knowledge of invariants
Diagnosis and recovery
Provide in platform
Domain context and vocabulary
Checks by default
Observable artifacts
Original scorecard derived from the report's limits
Measure expertise before the task
Review artifacts blind
Label trouble from telemetry
Check outcomes after adoption
People own
Goal and constraints
Meaningful verification
Decisions under trouble
Platform owns
Accessible context
Provable results
Boundaries and rollback
Verified-success gap narrows
Novices recover more often
Planning shifts to the agent
Quality remains stable
People still decide what
Competence captures the main gain
Expertise supports recovery
Outcomes beat action volume
Measure delegation bandwidth together with quality.
Not more activity. More verified work per unit of human judgment.
Book Cube
The primary report and appendix are linked on the first slide. Continue the discussion in Book Cube.
Alexander Polomodov, Technical Director & Fellow, T-Technologies
@Book_Cube