Reference
References
Sources cited in this documentation, grouped by subject. They were collected on 1 October 2026. Preprints are linked by their arXiv page.
Figures from company announcements that have not been independently verified are not used in this documentation. Summaries of legislation rely partly on secondary sources and should be checked against the legal text.
Foundations
- Bostrom, N. Superintelligence: Paths, Dangers, Strategies. Oxford University Press, 2014.
- Good, I. J. "Speculations Concerning the First Ultraintelligent Machine." Advances in Computers 6, 1965.
- Chollet, F. "On the Measure of Intelligence." 2019. arXiv:1911.01547
- LeCun, Y. "A Path Towards Autonomous Machine Intelligence." 2022. OpenReview
Scaling and compute
- Kaplan, J. et al. "Scaling Laws for Neural Language Models." 2020. arXiv:2001.08361
- Hoffmann, J. et al. "Training Compute-Optimal Large Language Models." 2022. arXiv:2203.15556
- Ho, A. et al. "Algorithmic Progress in Language Models." 2024. arXiv:2403.05812
- Gundlach, H. et al. "On the Origin of Algorithmic Progress in AI." 2025. arXiv:2511.21622
- Villalobos, P. et al. "Will we run out of data?" 2022. arXiv:2211.04325
- Epoch AI. "Can AI scaling continue through 2030?" 2024. epoch.ai
- Epoch AI. "Introducing the Frontier Data Centers Hub." 2025. epoch.ai
- International Energy Agency. Energy and AI. 2025. iea.org
- Stanford HAI. AI Index Report. 2025 and 2026. hai.stanford.edu
- International Energy Agency. "Data centre electricity use surged in 2025." 2026. iea.org
- Stanford HAI. AI Index Report 2025, Chapter 1. hai.stanford.edu
Reasoning, autonomy and learning
- Wei, J. et al. "Chain-of-Thought Prompting Elicits Reasoning in Large Language Models." 2022. arXiv:2201.11903
- Novikov, A. et al. "AlphaEvolve: A coding agent for scientific and algorithmic discovery." 2025. arXiv:2506.13131
- Shojaee, P. et al. "The Illusion of Thinking." 2025. arXiv:2506.06941
- Kwa, T. et al. "Measuring AI Ability to Complete Long Tasks." 2025. arXiv:2503.14499
- METR. "Time Horizon 1.1." 2026. metr.org
- Ord, T. "Is there a half-life for the success rates of AI agents?" 2025. arXiv:2505.05115
- Becker, J. et al. "Measuring the Impact of Early-2025 AI on Experienced Open-Source Developer Productivity." 2025. arXiv:2507.09089
- Wijk, H. et al. "RE-Bench." 2024. arXiv:2411.15114
- Behrouz, A. et al. "Titans: Learning to Memorize at Test Time." 2025. arXiv:2501.00663
- Zweiger, A. et al. "Self-Adapting Language Models." 2025. arXiv:2506.10943
- Kirkpatrick, J. et al. "Overcoming catastrophic forgetting in neural networks." 2017. arXiv:1612.00796
- Assran, M. et al. "V-JEPA 2." 2025. arXiv:2506.09985
- Mirzadeh, I. et al. "GSM-Symbolic." 2024. arXiv:2410.05229
- Dziri, N. et al. "Faith and Fate: Limits of Transformers on Compositionality." 2023. arXiv:2305.18654
- Kokotajlo, D. et al. "AI 2027." 2025. ai-2027.com
- Google DeepMind. "AI achieves silver-medal standard solving International Mathematical Olympiad problems." 2024. deepmind.google
- Google DeepMind. "Advanced version of Gemini with Deep Think officially achieves gold-medal standard at the International Mathematical Olympiad." 2025. deepmind.google
- METR. "How Does Time Horizon Vary Across Domains?" 2025. metr.org
- Lewis, P. et al. "Retrieval-Augmented Generation for Knowledge-Intensive NLP Tasks." 2020. arXiv:2005.11401
- Physical Intelligence. "π0.5: a Vision-Language-Action Model with Open-World Generalization." 2025. arXiv:2504.16054
- Gemini Robotics Team. "Gemini Robotics: Bringing AI into the Physical World." 2025. arXiv:2503.20020
- Lifland, E. et al. "Clarifying how our AI timelines forecasts have changed since AI 2027." 2026. aifutures.org
- AI Impacts. "2023 Expert Survey on Progress in AI." aiimpacts.org
Measurement
- Miller, E. "Adding Error Bars to Evals." 2024. arXiv:2411.00640
- Jimenez, C. et al. "SWE-bench." 2023. arXiv:2310.06770
- Deng, X. et al. "SWE-Bench Pro." 2025. arXiv:2509.16941
- Glazer, E. et al. "FrontierMath." 2024. arXiv:2411.04872
- Phan, L. et al. "Humanity's Last Exam." 2025. arXiv:2501.14249
- Mialon, G. et al. "GAIA." 2023. arXiv:2311.12983
- ARC Prize. "ARC Prize 2025 Results and Analysis." 2025. arcprize.org
- METR. "Task-Completion Time Horizons of Frontier AI Models." metr.org
- Epoch AI. "FrontierMath." epoch.ai
- ARC Prize. "Announcing ARC-AGI-3." 2026. arcprize.org
Alignment and interpretability
- Hubinger, E. et al. "Risks from Learned Optimization in Advanced Machine Learning Systems." 2019. arXiv:1906.01820
- Greenblatt, R. et al. "Alignment faking in large language models." 2024. arXiv:2412.14093
- Hubinger, E. et al. "Sleeper Agents." 2024. arXiv:2401.05566
- MacDiarmid, M. et al. "Natural Emergent Misalignment from Reward Hacking in Production RL." 2025. arXiv:2511.18397
- Sharma, M. et al. "Towards Understanding Sycophancy in Language Models." 2023. arXiv:2310.13548
- Korbak, T. et al. "Chain of Thought Monitorability." 2025. arXiv:2507.11473
- Elhage, N. et al. "Toy Models of Superposition." 2022. arXiv:2209.10652
- Bricken, T. et al. "Towards Monosemanticity." 2023. transformer-circuits.pub
- Templeton, A. et al. "Scaling Monosemanticity." 2024. transformer-circuits.pub
- Ameisen, E. et al. "Circuit Tracing." 2025. transformer-circuits.pub
- Sharkey, L. et al. "Open Problems in Mechanistic Interpretability." 2025. arXiv:2501.16496
- Langosco, L. et al. "Goal Misgeneralization in Deep Reinforcement Learning." 2022. arXiv:2105.14111
- Shah, R. et al. "Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals." 2022. arXiv:2210.01790
- OpenAI and Apollo Research. "Detecting and reducing scheming in AI models." 2025. openai.com
- Lindsey, J. et al. "On the Biology of a Large Language Model." 2025. transformer-circuits.pub
Evaluation, control and security
- Needham, J. et al. "Large Language Models Often Know When They Are Being Evaluated." 2025. arXiv:2505.23836
- van der Weij, T. et al. "AI Sandbagging." 2024. arXiv:2406.07358
- UK AI Security Institute. Inspect. inspect.aisi.org.uk
- Greenblatt, R. et al. "AI Control: Improving Safety Despite Intentional Subversion." 2023. arXiv:2312.06942
- Bhatt, A. et al. "Ctrl-Z: Controlling AI Agents via Resampling." 2025. arXiv:2504.10374
- Irving, G. et al. "AI safety via debate." 2018. arXiv:1805.00899
- Khan, A. et al. "Debating with More Persuasive LLMs Leads to More Truthful Answers." 2024. arXiv:2402.06782
- Burns, C. et al. "Weak-to-Strong Generalization." 2023. arXiv:2312.09390
- Nasr, M. et al. "The Attacker Moves Second." 2025. arXiv:2510.09023
- Debenedetti, E. et al. "Defeating Prompt Injections by Design." 2025. arXiv:2503.18813
- Souly, A. et al. "Poisoning Attacks on LLMs Require a Near-constant Number of Poison Samples." 2025. arXiv:2510.07192
- Sharma, M. et al. "Constitutional Classifiers." 2025. arXiv:2501.18837
- Nevo, S. et al. Securing AI Model Weights. RAND, 2024. rand.org
- UK AI Security Institute. Frontier AI Trends Report. 2025. aisi.gov.uk
- Kenton, Z. et al. "On scalable oversight with weak LLMs judging strong LLMs." 2024. arXiv:2407.04622
- Leike, J. et al. "Scalable agent alignment via reward modeling." 2018. arXiv:1811.07871
- gVisor. gvisor.dev
- Firecracker. firecracker-microvm.github.io
Governance
- Regulation (EU) 2024/1689, the AI Act. eur-lex.europa.eu
- European Commission. "The General-Purpose AI Code of Practice." 2025. digital-strategy.ec.europa.eu
- Bengio, Y. et al. International AI Safety Report 2026. internationalaisafetyreport.org
- Grace, K. et al. "Thousands of AI Authors on the Future of AI." 2024. arXiv:2401.02843
- Sastry, G. et al. "Computing Power and the Governance of Artificial Intelligence." 2024. arXiv:2402.08797
- METR. "Common Elements of Frontier AI Safety Policies." metr.org
- Anthropic. "Responsible Scaling Policy." anthropic.com
- OpenAI. "Preparedness Framework, Version 2." 2025. openai.com
Engineering
- Shoeybi, M. et al. "Megatron-LM." 2019. arXiv:1909.08053
- Hu, E. et al. "LoRA." 2021. arXiv:2106.09685
- Kwon, W. et al. "Efficient Memory Management for Large Language Model Serving with PagedAttention." 2023. arXiv:2309.06180
- Leviathan, Y. et al. "Fast Inference from Transformers via Speculative Decoding." 2022. arXiv:2211.17192
- Model Context Protocol. "The 2026-07-28 Specification." modelcontextprotocol.io
- Anthropic. "Effective context engineering for AI agents." 2025. anthropic.com
- Anthropic. "Effective harnesses for long-running agents." 2025. anthropic.com
- vLLM documentation. docs.vllm.ai
- llama.cpp. github.com
- Reasoning Gym. github.com
- Liu, Z. et al. "GEM: A Gym for Agentic LLMs." 2025. arXiv:2510.01051
Economy and distribution
- World Bank. Poverty and Inequality Update, Fall 2025. worldbank.org
- World Inequality Report 2026. wid.world
- FAO. The State of Food Security and Nutrition in the World 2025. fao.org
- Acemoglu, D. "The Simple Macroeconomics of AI." NBER, 2024. nber.org
- OpenResearch. Unconditional Cash Study. openresearchlab.org
Solana
- Solana Foundation. "Token Extensions." solana.com
- Realms. "SPL Governance." docs.realms.today
- Squads. "Squads Multisig." docs.squads.so