long-form

The strangest developer productivity metric of all time

Matthew Tyson argues token burn is a worse productivity measure than lines of code, pointing at Meta's Claudeonomics leaderboard, which ranked the top 250 of over 85,000 employees and drove 60.2 trillion tokens in 30 days.

Published 2026-08-12Source: InfoWorld
InfoWorld source artwork

Why it matters

Gamified token burn shows up as damage rather than output: Meta engineers tied site outages to careless generated code, Amazon killed its Kirorank board within weeks, and Uber reportedly drained a full year of AI budget by the end of Q1.

Tokenmaxxing read

The metric inverts value. An hour of LLM debugging that lands a three-line fix scores near zero, while 500 lines of generated spaghetti tops the board. Since output tokens can bill up to five times input, the dashboard actively rewards the most expensive habit.

Source takeaway

An opinion column, but the code-quality numbers are sourced: a June 2026 analysis of roughly 600 million code changes found duplication up 81%, refactoring down 70% versus 2022, and short-term churn doubling from 3.3% in 2021 to 7.1% in 2025.

Topic links

tokenmaxxingmetricstopicscoreboardsworkplace-aitopicai-roi
Related projects

Tools that match this angle

#2Direct
Observability

Langfuse

langfuse/langfuse

Open-source LLM engineering platform for observability, traces, metrics, evals, prompt management, datasets, and playground workflows.

33K3.6KSource-available
tracesevalscosts
#5Direct
Evaluation

promptfoo

promptfoo/promptfoo

A CLI and CI workflow for testing prompts, agents, and RAG systems across models, with evals and red-team style checks.

24.2K2.2KMIT
prompt-evalscirag
#6In spirit
Evaluation

DSPy

stanfordnlp/dspy

A framework for programming and optimizing language-model pipelines rather than hand-tuning one prompt at a time.

37.2K3.2KMIT
optimizationprogrammingevals
Related feed

More source-linked context

Generated Tokenmaxxing editorial thumbnail for ‘Tokenmaxxing’ Is the New Quiet Quitting—Here’s the Fix - SUCCESS Magazine
newsSM
news

‘Tokenmaxxing’ Is the New Quiet Quitting—Here’s the Fix - SUCCESS Magazine

SUCCESS argues tokenmaxxing-style adoption targets create performative AI usage. Their fix is to measure outcomes and quality, not raw token volume.

tokenmaxxingexplainerworkplace-ai
Read note
Generated Tokenmaxxing editorial thumbnail for Amazon employees admit to using AI unnecessarily to pump up internal usage scores — workers complain of intense pressure to use AI tools - Tom's Hardware
newsTH
news

Amazon employees admit to using AI unnecessarily to pump up internal usage scores — workers complain of intense pressure to use AI tools - Tom's Hardware

Amazon's internal AI usage targets can turn into tokenmaxxing: employees run unnecessary tasks in agent tools to climb dashboards rather than ship better work.

tokenmaxxingexplainerworkplace-ai
Read note
404 Media source artwork
news4M
news

Microsoft Tells Engineers ‘Tokenmaxxing Is Not What We Are Optimizing For’

Microsoft EVP Jay Parikh told staff that tokenmaxxing is not the goal, giving divisions AI token budget targets as of July 2026 and making OpenAI's cheaper GPT-5.6 the default model for internal use.

tokenmaxxingexplainerworkplace-ai
Read note