Ai token spend surged past 10x while delivery barely budged. plandek's q4 benchmark across 2,500…
ai token spend surged past 10x while delivery barely budged. plandek's q4 benchmark across 2,500 engineering teams, out oct 4, found the productivity promise still isn't showing up in the data.
the fastest-growing line item in software is the one with the weakest measured return.
Context
Plandek's press release, dated October 4, 2026 on its blog, announces its Q4 2026 AI Adoption and Impact Benchmarks, a study of quantitative data from more than 2,500 software engineering teams. It says AI token spend has increased 10 to 13 times since January 2025, and cites a Goldman Sachs forecast of a further 24 times increase in token consumption by 2030.
The release says the most mature agentic engineering teams raised normalized throughput by 50 percent over six months, while their Stories-to-Bugs ratio fell by 27 percent. Cycle time improved by 16 percent and Lead Time to Value by only 4 percent.
The 2,500 teams, the 10x spend and the October 4 date match Plandek's own release. This is a vendor study from a developer productivity platform, so its figures are Plandek's own and come from its customers' data, not an independent sample.
'Delivery barely budged' fits one measure. Lead Time to Value improved only 4 percent, but the same release reports a 50 percent throughput gain for the most mature teams and a 27 percent drop in the Stories-to-Bugs ratio, so the picture is throughput up, quality down and end-to-end delivery nearly flat.
The Goldman Sachs forecast is cited by Plandek and was not checked at its source. 'The weakest measured return' is the author's line.
Related work
- AI is not yet delivering on its productivity promise as token spend surges >10x (Plandek, October 4, 2026) ↗Source for the 2,500 teams, the 10 to 13 times token spend and the throughput, quality and lead time figures.
- AI Adoption and Impact Benchmarks Q4 2026 (Plandek) ↗Plandek's landing page for the Q4 2026 benchmarks.
Watch next
- Find an independent study of AI coding tool spend against delivery outcomes. Check the Goldman Sachs forecast at its source.
Sources
Provenance
The note above is reproduced unedited from the original post, first published on Threads on 9 October 2026 at 09:18 IST. Sources are the papers and datasets the note draws on.
View the original post ↗Embed this note
More notes
The air is now being asked to keep its own ledger
the air is now being asked to keep its own ledger: ecmwf’s aifs compo becomes the first ai model to forecast atmospheric composition globally every three hours, cleanair simulates 365 days of pm2.5 over china in ten seconds, and a unified framework maps six pollutants at one kilometer across the whole country. the air now files its own composition report.
read the note →The current is now being asked to draw its own map
the current is now being asked to draw its own map: china’s langya 2.0 predicts six ocean phenomena including internal waves and mesoscale eddies, a deep net called wenhai resolves eddies globally with air sea flux formulas built in, and scripps infers surface currents from the way temperature patterns deform in satellite images. the ocean now files its own circulation report.
read the note →The soil is now being asked to report its own carbon
the soil is now being asked to report its own carbon: a nix color sensor paired with generative data augmentation predicts soil organic carbon without a lab, random forest drives 74 percent of soil health mapping studies, and sentinel 2 tracks five year carbon change across france and italy from 922 samples. the dirt now files its own carbon account.
read the note →