The AI boom is quietly making your cloud bill bigger too

July 29, 2026

For about a decade, the cloud got cheaper almost every year — a steady drip of price cuts and new, cheaper instance types you could count on. That era is over. And the thing that ended it is the same thing running up your token bill: the AI buildout.

The evidence, from people who buy this stuff

A recent r/devops thread on cloud-vs-on-prem turned, unprompted, into a pile-up of the same observation from different people:

The decade of cloud getting cheaper every year is over.

Take that 3–5× as a Reddit guess, not a forecast — I won’t put a number on it. But the direction is not in doubt: cloud pricing pressure is up, for the first time in the platform era. If your five-year infrastructure plan quietly assumes “cloud keeps getting cheaper,” it’s built on an assumption that just stopped being true.

Why this is one problem, not two

Here’s the part that matters for anyone reading both halves of this site — the cloud posts and the AI posts. This trend reprices both meters upward at once:

  1. Your cloud meter costs more per unit. Every wasteful gigabyte of egress, every idle box, every over-provisioned instance doesn’t just get bigger as you grow — it gets more expensive per unit on top. The waste you shrugged off at old prices starts to bite. Controlling the meter isn’t a nice-to-have you can defer until “later, when there’s time” — later is more expensive than now.
  2. Your AI meter is the same story, sharper. The token prices you’re building on are venture-subsidized — below what it costs to serve you. As the buildout has to start paying for itself, that subsidy erodes: the uncapped meter gets more expensive per token and most teams still aren’t watching it. Rising unit price on an unwatched meter is exactly how a surprise bill gets bigger.

It’s one meter problem wearing two coats. The AI boom is the common cause, and it’s making both halves more urgent, together.

The move is the same on both meters: get off the part that’s about to reprice

You don’t fight a rising per-unit price by using slightly less of it. You get off the metered part where you can:

One honest caveat, because the same thread makes it: on-prem isn’t magically immune — hardware got more expensive too, so buying a server into a closet isn’t the escape. The escape is a flat-rate rented box (someone else owns and maintains the hardware; you pay a fixed price) for the steady, meter-heavy workloads — and hard caps on the AI spend. The point was never “flee to a closet.” It’s stop standing on the two meters that are about to go up.


If your cloud and AI bills are both riding meters that are about to reprice upward, moving the heavy parts to flat rate and capping the rest is exactly what I do — for both. Every message comes straight to me — I read and reply to each one myself, usually within a day, and what readers send shapes what I build next. It’s just me for now, so that’s genuinely true; it won’t be forever. Send me your setup and I’ll show you which lines are most exposed to the hikes — free, within a business day.

Free live workshop: cap your AI spend (Oct 8)

A hands-on 90-minute session — wire a fail-closed spend cap + cost-aware rate limiting against a real stack, live, so a leaked key or runaway agent can't run up your bill. Full details & times → Save your seat (and get each new lesson as it lands):

Double opt-in — one email to confirm. The lessons are free; the course is optional. No spam, unsubscribe anytime.