This post originally appeared in Exponential View.
βOlder models consume more compute per token.β
Cheaper AI was supposed to ease the compute crunch. Instead, it made it worse. The Jevons paradox, applied to intelligence, means that every time the price of a token falls, demand rises faster than supply can scale.
