Coding agents and productivity: what the METR studies actually show
A technical analysis of METR's measurements of AI-assisted software development, their limitations and the implications for credible internal evaluations.
The studies do not provide a universal productivity number for coding agents. They show how closely outcomes depend on tool maturity, task selection and measurement design.
Sources reviewed: 2026-07-18
Read the study →