ServiceNow warns tokenmaxxing can become a hype-cycle metric
The anti-vanity-metric case: buying more ingredients is not the same thing as running a better restaurant.
Published 2026-05-10Source: Observer
Why it matters
The ServiceNow angle is useful because it comes from the customer and enterprise-operations side of the market rather than from social chatter.
Tokenmaxxing read
It treats tokenmaxxing as a possible hype-cycle metric: impressive-looking consumption that can drift away from quality, governance, and actual workflow improvement.
Source takeaway
A strong source to cite when the site needs balance against the idea that more agent or model usage is automatically better.
Claude Code was using 51,000 tokens before I even typed a prompt — I fixed it
Mahnoor Faisal opened a new Claude Code session, ran /context, and found 51,400 tokens already loaded. Disabling four test plugins and auto-memory got the starting context down to roughly 41,400 before any real prompt.
Dropping Claude Code from High to Medium effort cut output tokens 45%
XDA's Mahnoor Faisal ran five coding jobs on Sonnet 5 twice from an identical starting codebase, changing only the effort level. High spent about 26,000 output tokens; Medium finished the same work on roughly 14,300.
The cost of intelligence: How CIOs can manage AI demand at scale - McKinsey & Company
McKinsey’s July 20 report finds 93% of enterprises are already blowing past their AI budgets, with spend jumping nearly 4x as pilots go company-wide. The fix it prescribes: run “FinOps for AI” and treat tokens like cloud cost.