news

Ramp Raises US$750m to Build Gen AI Infrastructure - AI Magazine

TechCrunch reports Ramp raised $750M at a $44B valuation, with CEO Eric Glyman casting cross-provider AI token-spend monitoring as Ramp's new 'third pillar' product.

Published 2026-06-11Source: TechCrunch
Generated Tokenmaxxing editorial thumbnail for Ramp Raises US$750m to Build Gen AI Infrastructure - AI Magazine

Why it matters

Spend-management is racing to own the token line item. Ramp, past $1B in annualized revenue with 70,000+ customers, is building cross-provider token monitoring and an agent corporate card, betting that controlling AI costs is the next fintech revenue stream.

Tokenmaxxing read

AI FinOps is going mainstream: the same week, Uber capped staff at $1,500 of AI spend after burning its 2026 budget in four months. When a $44B fintech treats AI tokens as a third pillar of corporate spend, token accounting becomes a line CFOs audit.

Source takeaway

TechCrunch is the original report; note its dig that Glyman's 'third pillar' blog reads AI-generated, and that the $1.5B run-rate figure comes via Bloomberg, not Ramp. Use it for the funding facts and the token-spend-as-product thesis.

Topic links

Related projects

Tools that match this angle

#4In spirit
Agents

LangGraph

langchain-ai/langgraph

A framework for building resilient stateful agents with explicit graphs, persistence, human-in-the-loop flows, and controllable execution.

39.9K6.7KMIT
agentsstateworkflows
#1Direct
Routing

LiteLLM

BerriAI/litellm

An OpenAI-compatible gateway and SDK for calling many model providers with budgets, logging, load balancing, guardrails, and cost tracking.

56.5K10.7KSource-available
gatewaycost-trackingrouting
#2Direct
Observability

Langfuse

langfuse/langfuse

Open-source LLM engineering platform for observability, traces, metrics, evals, prompt management, datasets, and playground workflows.

33.2K3.6KSource-available
tracesevalscosts
Related feed

More source-linked context

XDA source artwork
long-formX
long-form

Dropping Claude Code from High to Medium effort cut output tokens 45%

XDA's Mahnoor Faisal ran five coding jobs on Sonnet 5 twice from an identical starting codebase, changing only the effort level. High spent about 26,000 output tokens; Medium finished the same work on roughly 14,300.

coding-agentstoken-consumptionai-spend
Read note
theclimatebrink.com source artwork
newsT
news

The real energy use of agentic AI

Climate scientist Zeke Hausfather metered his own Claude Code habit: 1,138 typed prompts fanned out to more than 14,000 model calls and 3.2 billion tokens in eight weeks, drawing roughly 170 kWh of data-center electricity.

tokenmaxxingcoding-agentsagents
Read note
ZDNET source artwork
newsZ
news

Token-maxing is an AI cost sink - how to use agents without busting your budget

ZDNET asks enterprise leaders how to run agents without wrecking the budget. Boomi CEO Steve Lucas says he spent ten times more on Claude last year than the year before, and calls that pace flatly unsustainable.

tokenmaxxingagentstoken-consumption
Read note