TBPN

← Full issue

September 4, 2026

AI model costs may be better measured per completed task than per token

Token-pricing charts and benchmark comparisons are described as increasingly misleading because models differ in efficiency. A cited comparison says 3.8 Flash appears 13 times cheaper per token, while Astra is cheaper per completed task because it is more efficient.

The analysis argues that cost efficiency should be measured per task rather than per token. It is uncertain whether greater token efficiency would reduce overall token demand; token volumes may continue rising, but growth could decelerate if future models become dramatically more efficient, even as they perform more valuable work. Sundar Pichai has discussed exponential increases in token volumes at Google and Alphabet.

Privacy ·