Below you will find pages that utilize the taxonomy term “Margins”
Posts
Inference Prices Are Falling Faster Than Anyone's Budget Assumed
Microsoft priced its new speech recognition model at ten cents an audio hour. The previous model in the same line launched at thirty-six cents five months earlier. That’s a seventy-two percent cut in under half a year, on a product that also handles more languages and runs materially faster than it did.
Nothing about that is unusual any more, and the consistency is what makes it worth writing down. Per-token and per-unit prices across the major model providers have fallen by an order of magnitude or more since 2023, at every capability tier.