Back Issues/Search Home → Calendar → Archive → Current Issue → Popular →

All issuesVolume 341, Issue 3IT NewsAI

The Enterprise AI Cost Reckoning: Why Falling Per-Token Prices Aren't Saving You

AIwire, Wednesday, August 19th, 2026

AIwire explains why enterprise AI spend keeps climbing even as per-token inference prices collapse.

AIwire examines the gap between falling model prices and rising enterprise AI bills. Uber's CTO Praveen Neppalli Naga said in May that the budget earmarked for Claude Code was already blown, and Box CEO Aaron Levie warned that as agents move beyond engineering into legal and sales, compute budgets will rise monotonically.

Ramp's internal data shows enterprise token spend grew 13 times between January 2025 and early 2026. Meanwhile Andreessen Horowitz has documented LLMflation, in which inference costs for equivalent model performance fall 10 times a year, faster than Moore's Law. The article argues the volume of tokens consumed by agentic workflows is growing faster than unit prices are falling.

more →  ·  More from AI →