Canonical event

Chinchilla preprint reframes compute-optimal language-model scaling

DeepMind's March 29 preprint argued that many large language models were undertrained for their compute budgets and reported better performance by balancing model size with substantially more training tokens.

Artifacts

Signals

View canonical JSON