Issue 2026-09-23 · Industry · 研究 · 市场
AI tokens getting too cheap to meter
An analysis argues the cost of AI inference is falling by orders of magnitude each year with no sign of slowing. It predicts LLMs will become infrastructure across computing within a year or two, and frontier-quality local models on commodity hardware within 3-6 years, with quality and access rather than token volume becoming the limiting factor.
Read original ↗