tokens too cheap to meter

2026-09-16 • llms • economics

Table of contents

  1. Is this really happening?
    1. Improvements that affect all AI
    2. Improvements that affect hosted AI
    3. Improvements that affect local AI
    4. Improvements that affect specialized use cases
    5. Putting it together
  2. What happens next?
    1. Tokens become cheaper than tool calls
    2. Supply-side Jevons Paradox
    3. How are investors going to make their money back?
    4. Demand-side Jevons Paradox
    5. Optionality
  3. Summary

The price of using machine learning intelligence is decreasing by several orders of magnitude a year and shows no signs of slowing. We are likely to see LLMs integrated into every part of computing as infrastructure, not just as a product, in the next year or two. We are likely to see LLMs running locally at current frontier-quality on commodity hardware in the next 3-6 years. Starting very soon, we are likely to see quality and access become the limiting factor to AI 1 use, not sheer number of tokens.