Model Size Scaling in 2023-2031 (21 minute read)

TLDR AI News

Summary

An analysis of AI model size scaling trends from 2023 to 2031, published on LessWrong.

Targeting a particular speed of token generation puts a constraint on the total parameters of the model. If there isn't enough pretraining compute, models will remain smaller. This article looks at these considerations and estimates model sizes feasible for each year between 2023 and 2031. There are many assumptions that go into the estimates, which predict total parameters for models to reach 1.4 quadrillion in 2031.
Original Article
View Cached Full Text

Cached at: 06/23/26, 01:43 PM

# Model Size Scaling in 2023-2031 — LessWrong Source: [https://www.lesswrong.com/posts/yLHiQGCPdvzL9fBn3/model-size-scaling-in-2023-2031](https://www.lesswrong.com/posts/yLHiQGCPdvzL9fBn3/model-size-scaling-in-2023-2031) x Model Size Scaling in 2023\-2031 — LessWrong

Similar Articles

Scaling Laws, Carefully (25 minute read)

TLDR AI

A comprehensive overview of scaling laws in deep learning, tracing their theoretical roots and empirical findings, and explaining how loss decreases predictably with model size, data, and compute.

Why there is a lack of new 100B-120B models?

Reddit r/LocalLLaMA

Analysis of the trend in AI model sizes, noting a gap in the 100-120B parameter range with recent releases focusing on smaller (25-35B) or larger (200B+) models.

Scaling laws for neural language models

OpenAI Blog

Foundational empirical study demonstrating power-law scaling relationships between language model performance and model size, dataset size, and compute budget, with implications for optimal training allocation and sample efficiency.