The Intel Brief

Business · analysis

The quiet tax on every ‘AI feature’: your cloud bill

Inference is an operating expense that behaves like a headcount line.

The Intel Brief Desk · 28 Aug 2026

Key takeaways

  • Token volume is a forecastable cost if you instrument it.
  • Routing easy tasks to small models is an editorial pattern as well as an engineering one.
  • Cost per published article is a first-class metric in this newsroom.

Why this belongs on the business desk

AI features look like product. They spend like COGS. Each user action that calls a large model is closer to a taxi meter than to a one-time software licence.

What to watch

When a company says “AI-attached revenue,” ask whether they also disclose inference cost. The gap between those two numbers is where stories hide.