news

Ramp Raises US$750m to Build Gen AI Infrastructure - AI Magazine

TechCrunch reports Ramp raised $750M at a $44B valuation, with CEO Eric Glyman casting cross-provider AI token-spend monitoring as Ramp's new 'third pillar' product.

Published 2026-06-11Source: TechCrunch
Generated Tokenmaxxing editorial thumbnail for Ramp Raises US$750m to Build Gen AI Infrastructure - AI Magazine

Why it matters

Spend-management is racing to own the token line item. Ramp, past $1B in annualized revenue with 70,000+ customers, is building cross-provider token monitoring and an agent corporate card, betting that controlling AI costs is the next fintech revenue stream.

Tokenmaxxing read

AI FinOps is going mainstream: the same week, Uber capped staff at $1,500 of AI spend after burning its 2026 budget in four months. When a $44B fintech treats AI tokens as a third pillar of corporate spend, token accounting becomes a line CFOs audit.

Source takeaway

TechCrunch is the original report; note its dig that Glyman's 'third pillar' blog reads AI-generated, and that the $1.5B run-rate figure comes via Bloomberg, not Ramp. Use it for the funding facts and the token-spend-as-product thesis.

Topic links

Related projects

Tools that match this angle

#4In spirit
Agents

LangGraph

langchain-ai/langgraph

A framework for building resilient stateful agents with explicit graphs, persistence, human-in-the-loop flows, and controllable execution.

42.1K7.1KMIT
agentsstateworkflows
#1Direct
Routing

LiteLLM

BerriAI/litellm

An OpenAI-compatible gateway and SDK for calling many model providers with budgets, logging, load balancing, guardrails, and cost tracking.

59.3K11.6KSource-available
gatewaycost-trackingrouting
#2Direct
Observability

Langfuse

langfuse/langfuse

Open-source LLM engineering platform for observability, traces, metrics, evals, prompt management, datasets, and playground workflows.

34.9K3.8KSource-available
tracesevalscosts
Related feed

More source-linked context

XDA source artwork
long-formX
long-form

Claude Code was using 51,000 tokens before I even typed a prompt — I fixed it

Mahnoor Faisal opened a new Claude Code session, ran /context, and found 51,400 tokens already loaded. Disabling four test plugins and auto-memory got the starting context down to roughly 41,400 before any real prompt.

coding-agentstoken-consumptiontoken-waste
Read note
XDA source artwork
long-formX
long-form

Dropping Claude Code from High to Medium effort cut output tokens 45%

XDA's Mahnoor Faisal ran five coding jobs on Sonnet 5 twice from an identical starting codebase, changing only the effort level. High spent about 26,000 output tokens; Medium finished the same work on roughly 14,300.

coding-agentstoken-consumptionai-spend
Read note
theclimatebrink.com source artwork
newsT
news

The real energy use of agentic AI

Climate scientist Zeke Hausfather metered his own Claude Code habit: 1,138 typed prompts fanned out to more than 14,000 model calls and 3.2 billion tokens in eight weeks, drawing roughly 170 kWh of data-center electricity.

tokenmaxxingcoding-agentsagents
Read note