Urgent.News

What's breaking now, across thousands of outlets.

AI

AI token billing continues to cause sticker shock

Tech talk: are there inference providers that are transparent about token usage? A top pain point for companies using AI is the discrepancy between expected and actual invoice rates. Particularly when it comes to tokens. The increase in AI usage and tokenmaxxing has created a paradox where a provider knows exactly how many tokens it generated on your behalf, yet you find out only after the fact.…

Companies using AI are grappling with unexpected billing surprises related to token usage. Token usage has skyrocketed, leading to a paradox where providers know how many tokens they generate, but customers only learn after the bill arrives. This issue is particularly pronounced with reasoning models, where internal step-by-step reasoning is billed but remains hidden.

A 2025 study revealed that hidden reasoning tokens can make up over 90% of a model's total token spend for complex tasks. The situation has garnered industry attention, with two main responses emerging. Standards bodies like the Linux Foundation's proposed Tokenomics Foundation are working on open, vendor-neutral measurement standards for token accounting.

Meanwhile, providers such as Groq, Cerebras, and DigitalOcean are adopting transparent pricing models. DigitalOcean's Inference Router allows developers to route tasks to the most cost-effective models, providing budget control. As AI adoption accelerates, understanding token usage and its impact on billing is crucial for developers.

Written by urgent.news from Dev.to's reporting — not their text. Machine-written — may contain errors; check the original before relying on it.

Read the original at dev.to →

More in AI

More from Friday 18 September →