news

Beyond token-maxing: How US Bank AI chief navigates costs

U.S. Bank chief AI officer Prashant Mehrotra tells American Banker the bank never ran token maxing. Every proposed use case is bucketed by whether it drives growth, cuts cost, or merely feels productive — and the last bucket loses.

Published 2026-07-30Source: American Banker
American Banker source artwork

Why it matters

Mehrotra credits three forces for the retreat: model providers raising prices, an experimentation phase that has ended, and teams finally seeing enough value to operationalise it. A cost-discipline account from a regulated buyer, not a vendor.

Tokenmaxxing read

His filter is frequency, not sophistication: a dazzling use case that fires once a month loses to dull high-volume work with a measurable baseline. Experiments are timeboxed and killed when they miss, so the spending decision lands before any tokens burn.

Source takeaway

A Q&A, so every claim is the bank’s own framing with no independent verification and no spend figures. Useful for the operating model — evangelise, educate, enable, execute, and reuse over reinvention — not for benchmarking what U.S. Bank pays.

Topic links

Related projects

Tools that match this angle

#1Direct
Routing

LiteLLM

BerriAI/litellm

An OpenAI-compatible gateway and SDK for calling many model providers with budgets, logging, load balancing, guardrails, and cost tracking.

55.2K10.2KSource-available
gatewaycost-trackingrouting
#2Direct
Observability

Langfuse

langfuse/langfuse

Open-source LLM engineering platform for observability, traces, metrics, evals, prompt management, datasets, and playground workflows.

32.2K3.5KSource-available
tracesevalscosts
#10Direct
Routing

Portkey Gateway

Portkey-AI/gateway

An AI gateway for routing across LLMs with guardrails, provider abstraction, and an OpenAI-compatible API surface.

12.6K1.2KMIT
gatewayguardrailsrouting
Related feed

More source-linked context

ZDNET source artwork
newsZ
news

Token-maxing is an AI cost sink - how to use agents without busting your budget

ZDNET asks enterprise leaders how to run agents without wrecking the budget. Boomi CEO Steve Lucas says he spent ten times more on Claude last year than the year before, and calls that pace flatly unsustainable.

tokenmaxxingagentstoken-consumption
Read note
ContentGrip source artwork
newsC
news

Agencies confront rising AI costs

ContentGrip's Lena Marlowe reports agencies have moved past whether to use AI to whether they can prove it pays for its own token bill. Accountability, not adoption, is now the hard part.

tokenmaxxingagentstoken-consumption
Read note
Generated Tokenmaxxing editorial thumbnail for The cost of intelligence: How CIOs can manage AI demand at scale - McKinsey & Company
newsM&
news

The cost of intelligence: How CIOs can manage AI demand at scale - McKinsey & Company

McKinsey’s July 20 report finds 93% of enterprises are already blowing past their AI budgets, with spend jumping nearly 4x as pilots go company-wide. The fix it prescribes: run “FinOps for AI” and treat tokens like cloud cost.

tokenmaxxingfinopsai-spend
Read note