Inference Scaling and AI Governance

The shift from scaling up the pre-training compute of AI systems to scaling up their inference compute may have profound effects on AI governance. The nature of these effects depends crucially on whether this new inference compute will primarily be used during external deployment or as part of a more complex training programme within the lab. Rapid scaling of inference-at-deployment would: lower the importance of open-weight models (and of securing the weights of closed models), reduce the impact of the first human-level models, change the business model for frontier AI, reduce the need for power-intense data centres, and derail the current paradigm of AI governance via training compute thresholds. Rapid scaling of inference-during-training would have more ambiguous effects that range from a revitalisation of pre-training scaling to a form of recursive self-improvement via iterated distillation and amplification.

Paper

References (14)

06OpenAI, 12 Sep2024 · Learning to Reason with LLMs
07. Key Trends and Figures in Machine Learning2024 · epochai
08. Machines of Loving Grace: How AI Could Transform the World for the Better2024 · darioamodei.com
09model-free2017 · Benign model-free
102023. Trading Off Compute in Training and Inferenceepoch
112024. OpenAI and others seek new path to smarter AI as current methods hit limitations
122025. DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learningcs.CL

Scroll for more · 2 remaining

Similar papers

© 2026 NYSGPT2525 LLC