Race dynamics #
[Talking about times near the creation of the first AGI] you have the race dynamics where everyone's trying to stay ahead, and that might require compromising on safety. So I think you would probably need some coordination among the larger entities that are doing this kind of training [...] Pause either further training, or pause deployment, or avoiding certain types of training that we think might be riskier.
We already talked about race dynamics in the chapter on AI risks as amplifying factors for all risks. We mention them here again, because governance initiatives might have special leverage to be able to mitigate race dynamics.
Competition drives AI development at every level. From startups racing to demonstrate new capabilities to nation-states viewing AI leadership as essential for future power, competitive pressures shape how AI systems are built and deployed. This dynamic creates a prisoners dilemma like tension where even though everyone would benefit from careful, safety-focused development, those who move fastest gain competitive advantage (Hendryks, 2024).
The AI race creates a collective action problem. Even when developers recognize risks, unilateral caution means ceding ground to less scrupulous competitors. OpenAI's evolution illustrates this tension: founded as a safety-focused small nonprofit, competitive pressures led to creating a for-profit subsidiary and accelerating deployment timelines. When your competitors are raising billions and shipping products monthly, taking six extra months for safety testing feels like falling irreversibly behind (Gruetzemacher et al., 2024). This dynamic makes it exceedingly difficult for any single entity, be it a company or a country, to prioritize safety over speed (Askell et al., 2019).
Competitive pressure leads to safetywashing, cutting corners on testing, skipping external red-teaming, and rationalizing away warning signs. "Move fast and break things" becomes the implicit motto, even when the things being broken might include fundamental safety guarantees. We've already seen this with models released despite known vulnerabilities, justified by the need to maintain market position. Public companies face constant pressure to demonstrate progress to investors. Each competitor's breakthrough becomes an existential threat requiring immediate response. When Anthropic releases Claude 3, OpenAI must respond with GPT-4.5. When Google demonstrates new capabilities, everyone scrambles to match them. This quarter-by-quarter racing leaves little room for careful safety work that might take years to pay off.
National security concerns intensify race dynamics. When Vladimir Putin declared "whoever becomes the leader in AI will become the ruler of the world," he articulated what many policymakers privately believe (AP News, 2017). This transforms AI development from a commercial competition into a perceived struggle for geopolitical dominance. Over 50 countries have launched national AI strategies, often explicitly framing AI leadership as critical for economic and military superiority (Stanford HAI, 2024; Stanford HAI, 2025). Unlike corporate races measured in product cycles, international AI competition involves long-term strategic positioning. Yet paradoxically, this makes racing feel even more urgent: falling behind today might mean permanent disadvantage tomorrow.
Race dynamics make collective action and coordination feel impossible. Countries hesitate to implement strong safety regulations that might handicap their domestic AI industries. Companies resist voluntary safety commitments unless competitors make identical pledges. Everyone waits for others to move first, creating gridlock even when all parties privately acknowledge the risks. The result is a lowest-common-denominator approach to safety that satisfies no one.
AI governance needs innovative approaches to break out of race dynamics. Traditional arms control offers limited lessons, since AI development happens in private companies, not government labs. We need new approaches (Trajano & Ang, 2023; Barnett, 2025). Several ideas have been proposed. Some examples are:
- Reciprocal snap-back limits. States publish caps on model scale, autonomous-weapon deployment and data-center compute that activate only when peers file matching commitments. The symmetry removes the fear of unilateral restraint and keeps incentives focused on shared security rather than zero-sum dominance (Karnofsky, 2024).
- Safety as a competitive asset. Labs earn market trust by subjecting frontier models to independent red-team audits, embedding provenance watermarks and disclosing incident reports. Regulation can turn these practices into a de-facto licence to operate so that “secure by design” becomes the shortest route to sales (Shevlane et al., 2023; Tamirisa et al., 2024).
- Containment. Export controls on advanced chips; API-only access with real-time misuse monitoring; digital forensics; and Know-Your-Customer checks slow the spread of dangerous capabilities even as beneficial services stay widely available. These measures address open publication, model theft, talent mobility and hardware diffusion; factors that let a single leak replicate worldwide within days (Shevlane et al., 2023; Seger, 2023; Nevo et al., 2024).
- Agile multilateral oversight with a coordinated halt option. A lean UN-mandated body (think about it like a CERN or an IAEA-for-AI) needs the authority to impose emergency pauses when red-lines are crossed, backed by chip export restrictions and cloud-provider throttles that make a global “off switch” technically credible (Karnofsky, 2024; Petropoulos et al., 2025).
- Secret-safe verification. Secure enclaves, tamper-evident compute logs and zero-knowledge proofs let inspectors confirm that firms observe model and data controls without exposing weights or proprietary code, closing the principal oversight gap identified in current treaty proposals (Shevlane et al., 2023; Wasil et al., 2024; Anderljung et al. 2024).
Proliferation #
AI capabilities propagate globally through digital networks at speeds that render traditional control mechanisms largely ineffective. Unlike nuclear weapons that require specialized materials and facilities, AI models are patterns of numbers that can be copied and transmitted instantly. Let's think about this scenario - a cutting-edge AI model, capable of generating hyper-realistic deepfakes or designing novel bioweapons, is developed by a well-intentioned research lab. The lab, adhering to principles of open science, publishes their findings and releases the model's code as open-source. Within hours, the model is downloaded thousands of times across the globe. Within days, modified versions start appearing on code-sharing platforms. Within weeks, the capabilities that were once confined to a single lab have proliferated across the internet, accessible to anyone with a decent computer and an internet connection. This scenario, while hypothetical, isn't far from reality. This fundamental difference makes traditional non-proliferation approaches nearly useless for AI governance.
Multiple channels enable rapid proliferation:
- Open publication accelerates capability diffusion. The AI research community's commitment to openness means breakthrough techniques often appear on arXiv within days of discovery. What took one lab years to develop can be replicated by others in months. Meta's release of Llama 2 led to thousands of fine-tuned variants within weeks, including versions with safety features removed and new dangerous capabilities added (Seger, 2023).
- Model theft presents growing risks. As AI models become more valuable, they become attractive targets for malicious hackers and criminal groups. A single successful breach could transfer capabilities worth billions in development costs. Even without direct theft, techniques like model distillation can extract capabilities from API access alone (Nevo et al., 2024).
- Talent mobility spreads tacit knowledge. When researchers move between organizations, they carry irreplaceable expertise. The deep learning diaspora from Google Brain and DeepMind seeded AI capabilities worldwide. Unlike written knowledge, this experiential understanding of how to build and train models can't be controlled through traditional means (Besiroglu, 2024).
- Hardware proliferation enables distributed development. As AI chips become cheaper and more available, the barrier to entry keeps dropping. What required a supercomputer in 2018 now runs on hardware costing under 100,000 dollars. This democratization means dangerous capabilities become accessible to ever-smaller actors (Masi, 2024).
AI proliferation poses unique challenges - digital goods follow different rules than physical objects. Traditional proliferation controls assume scarcity: there's only so much enriched uranium or only so many advanced missiles. But copying a model file costs essentially nothing. Once capabilities exist anywhere, preventing their spread becomes a battle against the fundamental nature of information. It's far easier to share a model than to prevent its spread. Even sophisticated watermarking or encryption schemes can be defeated by determined actors.
Verifying that someone if not developing harmful AI capabilities is extremely hard. Unlike nuclear technology where detection capabilities roughly match proliferation methods, AI governance lacks comparable defensive tools (Shevlane, 2024). Nuclear inspectors can use satellites and radiation detectors to monitor compliance. But verifying that an organization isn't developing dangerous AI capabilities would require invasive access to code, data and development: practices likely revealing valuable intellectual property. Many organizations thus refuse intrusive monitoring (Wasil et al., 2024). This would require a combination of many different technical, and national measures.
Dual-use nature complicates controls. The same transformer architecture that powers beneficial applications can also enable harmful uses. Unlike specialized military technology, we can't simply ban dangerous AI capabilities without eliminating beneficial ones. This dual-use problem means governance must be far more nuanced than traditional non-proliferation regimes (Anderljung, 2024). A motivated individual with modest resources can now fine-tune powerful models for harmful purposes. This democratization of capabilities means threats can emerge from anywhere, not just nation-states or major corporations. Traditional governance frameworks aren't designed for this level of distributed risk.
How can governance help slow AI proliferation? Several potential solutions have been proposed to find the right balance between openness and control:
- Targeted openness. Publish fundamental research but withhold model weights and fine-tuning recipes for high-risk capabilities, keeping collaboration alive while denying turnkey misuse (Seger, 2023).
- Staged releases. Roll out progressively stronger versions only after each tier passes red-team audits and external review, giving society time to surface failure modes and tighten safeguards before the next step (Solaiman, 2023).
- Enhanced information security. Treat frontier checkpoints like crown-jewel secrets: hardened build pipelines, model-weight encryption in use and at rest, and continuous insider-threat monitoring (Nevo et al., 2024).
- Export controls and compute access restrictions. Block shipment of the most advanced AI accelerators to unvetted end-users and require cloud providers to gate high-end training clusters behind Know-Your-Customer checks (O’Brien et al., 2024).
- Responsible disclosure. Adopt cybersecurity-style norms for reporting newly discovered “dangerous capability routes,” so labs alert peers and regulators without publishing full exploit paths (O’Brien et al., 2024).
- Built-in technical brakes. Embed jailbreak-resistant tuning, capability throttles and provenance watermarks that survive model distillation, adding friction even after weights leak (Dong et al., 2024).
Uncertainty #
The exact way the post-AGI world will look is hard to predict — that world will likely be more different from today's world than today's is from the 1500s [...] We do not yet know how hard it will be to make sure AGIs act according to the values of their operators. Some people believe it will be easy; some people believe it'll be unimaginably difficult; but no one knows for sure.
Expert predictions consistently fail to capture AI's actual trajectory. If you read media coverage of ChatGPT — which called it ‘breathtaking’, ‘dazzling’, ‘astounding’ — you’d get the sense that large language models (LLMs) took the world completely by surprise. Is that impression accurate? Actually, yes. (Cotra, 2023) GPT-3's capabilities exceeded what many thought possible with simple scaling. Each major breakthrough seems to come from unexpected directions, making long-term planning nearly impossible (Gruetzemacher et al., 2021; Grace et al., 2017). The "scaling hypothesis" (larger models with more compute reliably produce more capable systems) has held surprisingly well. But we don't know if this continues to AGI or hits fundamental technical or economic limits. This uncertainty has massive governance implications. If scaling continues, compute controls remain effective. If algorithmic breakthroughs matter more, entirely different governance approaches are needed (Patel, 2023).
Risk assessments vary by orders of magnitude. Some researchers assign negligible probability to existential risks from AI, while others consider them near-certain without intervention, reflecting fundamental uncertainty about AI's trajectory and controllability. When experts disagree this dramatically, how can policymakers make informed decisions? (Narayanan & Kapoor, 2024).
Capability emergence surprises even developers. Models demonstrate abilities their creators didn't anticipate and can't fully explain (Cotra, 2023). If the people building these systems can't predict their capabilities, how can governance frameworks anticipate what needs regulating? This unpredictability compounds with each generation of more powerful models (Grace et al., 2024). Traditional policy-making assumes predictable outcomes. Environmental regulations model pollution impacts. Drug approval evaluates specific health effects. But AI governance must prepare for scenarios ranging from gradual capability improvements to sudden recursive self-improvement.
Waiting for certainty means waiting too long. By the time we know exactly what AI capabilities will emerge, it may be too late to govern them effectively. Yet acting under uncertainty risks implementing wrong-headed policies that stifle beneficial development or fail to prevent actual risks. This creates a debilitating dilemma for conscientious policymakers (Casper, 2024).
How can governance operate under uncertainty? Adaptive governance models that could keep pace with rapidly changing technology could offer a path forward. Rather than fixed regulations based on current understanding, we need frameworks that can evolve with our knowledge. This might include:
- Regulatory triggers based on capability milestones rather than timelines
- Sunset clauses that force regular reconsideration of rules
- Safe harbors for experimentation within controlled environments
- Rapid-response institutions capable of updating policies as understanding improves
Building consensus despite uncertainty requires new approaches. Traditional policy consensus emerges from shared understanding of problems and solutions. With AI, we lack both. Yet somehow we must build sufficient agreement to implement governance before capabilities outrace our ability to control them. This may require focusing on process legitimacy rather than outcome certainty agreeing on how to make decisions even when we disagree on what to decide.
Accountability #
[After resigning from OpenAI] These problems are quite hard to get right, and I am concerned we aren't on a trajectory to get there [...] OpenAI is shouldering an enormous responsibility on behalf of all of humanity. But over the past years, safety culture and processes have taken a backseat to shiny products. We are long overdue in getting incredibly serious about the implications of AGI.
A small number of actors make decisions that affect all of humanity. The CEOs of perhaps five companies and key officials in three governments largely determine how frontier AI develops. Their choices about what to build, when to deploy, and how to ensure safety have consequences for billions who have no voice in these decisions. OpenAI's board has fewer than ten members. Anthropic's Long-Term Benefit Trust controls the company with just five trustees. These tiny groups make decisions about technologies that could fundamentally alter human society. No pharmaceutical company could release a new drug with such limited oversight, yet AI systems with far broader impacts face minimal external scrutiny. Nearly all frontier AI development happens in just two regions: the San Francisco Bay Area and London. The values, assumptions, and blind spots of these tech hubs shape AI systems used worldwide, yet we know more about how sausages are made than how frontier AI systems are trained. What seems obvious in Palo Alto might be alien in Lagos or Jakarta, yet the global majority have essentially no input into AI development (Adan et al., 2024).
Traditional accountability mechanisms don't apply. Corporate boards nominally provide oversight, but most lack the incentives to evaluate systemic AI risks. Government regulators struggle to keep pace with rapid development. Academic researchers who might provide scientific evidence and independent assessment often depend on corporate funding or compute access. The result is a governance vacuum where no one has both the capability and authority needed for proper governance (Anderljung, 2023). The consequences of this lack of governance are already becoming apparent. We've seen AI-generated deepfakes used to spread political misinformation (Swenson & Chan, 2024). Language models have been used to create convincing phishing emails and other scams (Stacey, 2025). When models demonstrate concerning behaviors, we can't trace whether they result from training data , reward functions, or architectural choices. This black box nature of development is a big bottleneck in accountability (Chan et al., 2024).
Power and Wealth Concentration #
AI concentrates power in unprecedented ways. AI systems, especially those developed by dominant corporations, are reshaping societal power structures. These systems determine access to information and resources, effectively exercising automated authority over individuals (Lazar, 2024). As these systems become more capable, this concentration intensifies. The organization that first develops AGI could gain decisive advantages across every domain of human activity, a winner-take-all dynamic with no historical precedent.
Wealth effects compound existing inequalities. AI automation primarily benefits capital owners while displacing workers, deepening existing disparities. Recent empirical evidence suggests that AI adoption significantly increases wealth inequality by disproportionately benefiting those who own models, data, and computational resources, at the expense of labor (Skare et al., 2024). Without targeted governance interventions, AI risks creating never before seen levels of economic inequality, potentially resulting in the most unequal society in human history (O’Keefe, 2020).
Democratic governance faces existential challenges. When information itself is controlled by private entities, traditional democratic institutions struggle to remain effective (Kreps & Kriner, 2023). Some empirical evidence indicates that higher levels of AI integration correlate with declining democratic participation and accountability, as elected officials find themselves unable to regulate complex technologies that evolve faster than legislative processes (Chehoudi, 2025). This emerging technocratic reality fundamentally undermines democratic principles regarding public control and oversight.
International disparities threaten global stability. Countries without domestic AI capabilities face permanent subordination to AI leaders. AI adoption significantly exacerbates international inequalities, disproportionately favoring technologically advanced nations. This disparity threatens not only economic competitiveness but also basic sovereignty when critical decisions are effectively outsourced to foreign-controlled AI systems (Cerutti et al., 2025). We have no agreed frameworks for distributing AI's benefits or managing its disruptions. Should AI developers owe obligations to displaced workers? How should AI-generated wealth be taxed and redistributed? What claims do non-developers have on AI capabilities? These questions need answers before AI's impacts become irreversible, yet governance current discussions barely acknowledge them (Ding & Dafoe, 2024).
References
- Adan, S. N. et al. (2024). Voice and Access in AI: Global AI Majority Participation in Artificial Intelligence Development and Governance.Adan, S. N., Trager, R., Blomquist, K., Dennis, C., Edom, G., Velasco, L., Abungu, C., Garfinkel, B., Jacobs, J., Okolo, C. T., Wu, B., & Vipra, J. (2024). Voice and Access in AI: Global AI Majority Participation in Artificial Intelligence Development and Governance. Oxford Martin AI Governance Initiative, University of Oxford. https://oms-www.files.svdcdn.com/production/downloads/reports/Voice%20and%20Access%20in%20AI_%20Global%20AI%20Majority%20Participation%20in%20Artificial%20Intelligence%20Development%20and%20Governance-%20final.pdf?dm=1729247034Adan, S. N., R. Trager, K. Blomquist, et al. 2024. Voice and Access in AI: Global AI Majority Participation in Artificial Intelligence Development and Governance. Oxford Martin AI Governance Initiative, University of Oxford. https://oms-www.files.svdcdn.com/production/downloads/reports/Voice%20and%20Access%20in%20AI_%20Global%20AI%20Majority%20Participation%20in%20Artificial%20Intelligence%20Development%20and%20Governance-%20final.pdf?dm=1729247034.Adan, S. N., et al. Voice and Access in AI: Global AI Majority Participation in Artificial Intelligence Development and Governance. Oxford Martin AI Governance Initiative, University of Oxford, Oct. 2024, https://oms-www.files.svdcdn.com/production/downloads/reports/Voice%20and%20Access%20in%20AI_%20Global%20AI%20Majority%20Participation%20in%20Artificial%20Intelligence%20Development%20and%20Governance-%20final.pdf?dm=1729247034.Adan, S. N. et al. Voice and Access in AI: Global AI Majority Participation in Artificial Intelligence Development and Governance. https://oms-www.files.svdcdn.com/production/downloads/reports/Voice%20and%20Access%20in%20AI_%20Global%20AI%20Majority%20Participation%20in%20Artificial%20Intelligence%20Development%20and%20Governance-%20final.pdf?dm=1729247034 (2024).S. N. Adan et al., “Voice and Access in AI: Global AI Majority Participation in Artificial Intelligence Development and Governance”, Oxford Martin AI Governance Initiative, University of Oxford, Oct. 2024. [Online]. Available: https://oms-www.files.svdcdn.com/production/downloads/reports/Voice%20and%20Access%20in%20AI_%20Global%20AI%20Majority%20Participation%20in%20Artificial%20Intelligence%20Development%20and%20Governance-%20final.pdf?dm=1729247034
- Anderljung, M. & Hazell, J. (2023). Protecting Society from AI Misuse: When are Restrictions on Capabilities Warranted?. arXiv.Anderljung, M., & Hazell, J. (2023). Protecting Society from AI Misuse: When are Restrictions on Capabilities Warranted?. In arXiv. https://arxiv.org/abs/2303.09377Anderljung, M., and J. Hazell. 2023. “Protecting Society from AI Misuse: When Are Restrictions on Capabilities Warranted?”. In arXiv. Preprint, March 16. https://arxiv.org/abs/2303.09377.Anderljung, M., and J. Hazell. “Protecting Society from AI Misuse: When Are Restrictions on Capabilities Warranted?”. arXiv, 16 Mar. 2023, https://arxiv.org/abs/2303.09377.Anderljung, M. & Hazell, J. Protecting Society from AI Misuse: When are Restrictions on Capabilities Warranted?. arXiv Preprint at https://arxiv.org/abs/2303.09377 (2023).M. Anderljung and J. Hazell, “Protecting Society from AI Misuse: When are Restrictions on Capabilities Warranted?”, Mar. 16, 2023. [Online]. Available: https://arxiv.org/abs/2303.09377
- Anderljung, M. et al. (2023). Towards Publicly Accountable Frontier LLMs: Building an External Scrutiny Ecosystem under the ASPIRE Framework. arXiv.Anderljung, M., Smith, E. T., O'Brien, J., Soder, L., Bucknall, B., Bluemke, E., Schuett, J., Trager, R., Strahm, L., & Chowdhury, R. (2023). Towards Publicly Accountable Frontier LLMs: Building an External Scrutiny Ecosystem under the ASPIRE Framework. In arXiv. https://arxiv.org/abs/2311.14711Anderljung, M., E. T. Smith, J. O'Brien, et al. 2023. “Towards Publicly Accountable Frontier LLMs: Building an External Scrutiny Ecosystem Under the ASPIRE Framework”. In arXiv. Preprint, November 15. https://arxiv.org/abs/2311.14711.Anderljung, M., et al. “Towards Publicly Accountable Frontier LLMs: Building an External Scrutiny Ecosystem Under the ASPIRE Framework”. arXiv, 15 Nov. 2023, https://arxiv.org/abs/2311.14711.Anderljung, M. et al. Towards Publicly Accountable Frontier LLMs: Building an External Scrutiny Ecosystem under the ASPIRE Framework. arXiv Preprint at https://arxiv.org/abs/2311.14711 (2023).M. Anderljung et al., “Towards Publicly Accountable Frontier LLMs: Building an External Scrutiny Ecosystem under the ASPIRE Framework”, Nov. 15, 2023. [Online]. Available: https://arxiv.org/abs/2311.14711
- AP News (2017). Putin: Leader in artificial intelligence will rule world. AP News.AP News. (2017, September 1). Putin: Leader in artificial intelligence will rule world. AP News. https://apnews.com/article/bb5628f2a7424a10b3e38b07f4eb90d4AP News. 2017. “Putin: Leader in Artificial Intelligence Will Rule World”. AP News, September 1. https://apnews.com/article/bb5628f2a7424a10b3e38b07f4eb90d4.AP News. “Putin: Leader in Artificial Intelligence Will Rule World”. AP News, 1 Sept. 2017, https://apnews.com/article/bb5628f2a7424a10b3e38b07f4eb90d4.AP News. Putin: Leader in artificial intelligence will rule world. AP News https://apnews.com/article/bb5628f2a7424a10b3e38b07f4eb90d4 (2017).AP News, “Putin: Leader in artificial intelligence will rule world”, AP News. [Online]. Available: https://apnews.com/article/bb5628f2a7424a10b3e38b07f4eb90d4
- Arvind Narayanan & Sayash Kapoor (2024). AI existential risk probabilities are too unreliable to inform policy.Arvind Narayanan, & Sayash Kapoor. (2024, July 26). AI existential risk probabilities are too unreliable to inform policy. https://aisnakeoil.com/p/ai-existential-risk-probabilitiesArvind Narayanan, and Sayash Kapoor. 2024. AI Existential Risk Probabilities Are Too Unreliable to Inform Policy. Edition. July 26. https://aisnakeoil.com/p/ai-existential-risk-probabilities.Arvind Narayanan, and Sayash Kapoor. AI Existential Risk Probabilities Are Too Unreliable to Inform Policy. 26 July 2024, https://aisnakeoil.com/p/ai-existential-risk-probabilities.Arvind Narayanan & Sayash Kapoor. AI existential risk probabilities are too unreliable to inform policy. https://aisnakeoil.com/p/ai-existential-risk-probabilities (2024).Arvind Narayanan and Sayash Kapoor, “AI existential risk probabilities are too unreliable to inform policy”. [Online]. Available: https://aisnakeoil.com/p/ai-existential-risk-probabilities
- Askell, A., Brundage, M. & Hadfield, G. (2019). The Role of Cooperation in Responsible AI Development. arXiv.Askell, A., Brundage, M., & Hadfield, G. (2019). The Role of Cooperation in Responsible AI Development. In arXiv. https://arxiv.org/abs/1907.04534Askell, A., M. Brundage, and G. Hadfield. 2019. “The Role of Cooperation in Responsible AI Development”. In arXiv. Preprint, July 10. https://arxiv.org/abs/1907.04534.Askell, A., et al. “The Role of Cooperation in Responsible AI Development”. arXiv, 10 July 2019, https://arxiv.org/abs/1907.04534.Askell, A., Brundage, M. & Hadfield, G. The Role of Cooperation in Responsible AI Development. arXiv Preprint at https://arxiv.org/abs/1907.04534 (2019).A. Askell, M. Brundage, and G. Hadfield, “The Role of Cooperation in Responsible AI Development”, Jul. 10, 2019. [Online]. Available: https://arxiv.org/abs/1907.04534
- Barnett, P. & Scher, A. (2025). AI Governance to Avoid Extinction: The Strategic Landscape and Actionable Research Questions.Barnett, P., & Scher, A. (2025). AI Governance to Avoid Extinction: The Strategic Landscape and Actionable Research Questions. Machine Intelligence Research Institute. https://techgov.intelligence.org/research/ai-governance-to-avoid-extinctionBarnett, P., and A. Scher. 2025. AI Governance to Avoid Extinction: The Strategic Landscape and Actionable Research Questions. Machine Intelligence Research Institute. https://techgov.intelligence.org/research/ai-governance-to-avoid-extinction.Barnett, P., and A. Scher. AI Governance to Avoid Extinction: The Strategic Landscape and Actionable Research Questions. Machine Intelligence Research Institute, May 2025, https://techgov.intelligence.org/research/ai-governance-to-avoid-extinction.Barnett, P. & Scher, A. AI Governance to Avoid Extinction: The Strategic Landscape and Actionable Research Questions. https://techgov.intelligence.org/research/ai-governance-to-avoid-extinction (2025).P. Barnett and A. Scher, “AI Governance to Avoid Extinction: The Strategic Landscape and Actionable Research Questions”, Machine Intelligence Research Institute, May 2025. [Online]. Available: https://techgov.intelligence.org/research/ai-governance-to-avoid-extinction
- Besiroglu, T., Bergerson, S. A., Michael, A., Heim, L., Luo, X. & Thompson, N. (2024). The Compute Divide in Machine Learning: A Threat to Academic Contribution and Scrutiny?. arXiv.Besiroglu, T., Bergerson, S. A., Michael, A., Heim, L., Luo, X., & Thompson, N. (2024). The Compute Divide in Machine Learning: A Threat to Academic Contribution and Scrutiny?. In arXiv. https://arxiv.org/abs/2401.02452Besiroglu, T., S. A. Bergerson, A. Michael, L. Heim, X. Luo, and N. Thompson. 2024. “The Compute Divide in Machine Learning: A Threat to Academic Contribution and Scrutiny?”. In arXiv. Preprint, January 4. https://arxiv.org/abs/2401.02452.Besiroglu, T., et al. “The Compute Divide in Machine Learning: A Threat to Academic Contribution and Scrutiny?”. arXiv, 4 Jan. 2024, https://arxiv.org/abs/2401.02452.Besiroglu, T. et al. The Compute Divide in Machine Learning: A Threat to Academic Contribution and Scrutiny?. arXiv Preprint at https://arxiv.org/abs/2401.02452 (2024).T. Besiroglu, S. A. Bergerson, A. Michael, L. Heim, X. Luo, and N. Thompson, “The Compute Divide in Machine Learning: A Threat to Academic Contribution and Scrutiny?”, Jan. 04, 2024. [Online]. Available: https://arxiv.org/abs/2401.02452
- Casper, S., Krueger, D. & Hadfield-Menell, D. (2025). Pitfalls of Evidence-Based AI Policy. arXiv.Casper, S., Krueger, D., & Hadfield-Menell, D. (2025). Pitfalls of Evidence-Based AI Policy. In arXiv. https://arxiv.org/abs/2502.09618Casper, S., D. Krueger, and D. Hadfield-Menell. 2025. “Pitfalls of Evidence-Based AI Policy”. In arXiv. Preprint, February 13. https://arxiv.org/abs/2502.09618.Casper, S., et al. “Pitfalls of Evidence-Based AI Policy”. arXiv, 13 Feb. 2025, https://arxiv.org/abs/2502.09618.Casper, S., Krueger, D. & Hadfield-Menell, D. Pitfalls of Evidence-Based AI Policy. arXiv Preprint at https://arxiv.org/abs/2502.09618 (2025).S. Casper, D. Krueger, and D. Hadfield-Menell, “Pitfalls of Evidence-Based AI Policy”, Feb. 13, 2025. [Online]. Available: https://arxiv.org/abs/2502.09618
- Cerutti et al. (2025). The Global Impact of AI – Mind the Gap. IMF eLibrary.Cerutti et al. (2025). The Global Impact of AI – Mind the Gap. Internet Archive (https://web.archive.org/web/20260909150837/https://www.elibrary.imf.org/view/journals/001/2025/076/article-A001-en.xml). IMF eLibrary. https://elibrary.imf.org/view/journals/001/2025/076/article-A001-en.xmlCerutti et al. 2025. “The Global Impact of AI – Mind the Gap”. IMF eLibrary. Https://web.archive.org/web/20260909150837/https://www.elibrary.imf.org/view/journals/001/2025/076/article-A001-en.xml. Internet Archive. https://elibrary.imf.org/view/journals/001/2025/076/article-A001-en.xml.Cerutti et al. “The Global Impact of AI – Mind the Gap”. IMF eLibrary, 2025, Internet Archive, https://web.archive.org/web/20260909150837/https://www.elibrary.imf.org/view/journals/001/2025/076/article-A001-en.xml, https://elibrary.imf.org/view/journals/001/2025/076/article-A001-en.xml.Cerutti et al. The Global Impact of AI – Mind the Gap. IMF eLibrary https://elibrary.imf.org/view/journals/001/2025/076/article-A001-en.xml (2025).Cerutti et al., “The Global Impact of AI – Mind the Gap”, IMF eLibrary. Accessed: Sep. 09, 2026. [Online]. Available: https://elibrary.imf.org/view/journals/001/2025/076/article-A001-en.xml
- Chan, A. et al. (2024). Visibility into AI Agents. arXiv.Chan, A., Ezell, C., Kaufmann, M., Wei, K., Hammond, L., Bradley, H., Bluemke, E., Rajkumar, N., Krueger, D., Kolt, N., Heim, L., & Anderljung, M. (2024). Visibility into AI Agents. In arXiv. https://arxiv.org/abs/2401.13138Chan, A., C. Ezell, M. Kaufmann, et al. 2024. “Visibility into AI Agents”. In arXiv. Preprint, January 23. https://arxiv.org/abs/2401.13138.Chan, A., et al. “Visibility into AI Agents”. arXiv, 23 Jan. 2024, https://arxiv.org/abs/2401.13138.Chan, A. et al. Visibility into AI Agents. arXiv Preprint at https://arxiv.org/abs/2401.13138 (2024).A. Chan et al., “Visibility into AI Agents”, Jan. 23, 2024. [Online]. Available: https://arxiv.org/abs/2401.13138
- Chehoudi, R. (2025). Artificial intelligence and democracy: pathway to progress or decline?. Journal of Information Technology & Politics.Chehoudi, R. (2025). Artificial intelligence and democracy: pathway to progress or decline?. Journal of Information Technology & Politics. https://doi.org/10.1080/19331681.2025.2473994Chehoudi, R. 2025. “Artificial Intelligence and Democracy: Pathway to Progress or Decline?”. Journal of Information Technology & Politics, ahead of print, March 6. https://doi.org/10.1080/19331681.2025.2473994.Chehoudi, R. “Artificial Intelligence and Democracy: Pathway to Progress or Decline?”. Journal of Information Technology & Politics, Mar. 2025, https://doi.org/10.1080/19331681.2025.2473994.Chehoudi, R. Artificial intelligence and democracy: pathway to progress or decline?. Journal of Information Technology & Politics https://doi.org/10.1080/19331681.2025.2473994 (2025) doi:10.1080/19331681.2025.2473994.R. Chehoudi, “Artificial intelligence and democracy: pathway to progress or decline?”, Journal of Information Technology & Politics, Mar. 2025, doi: 10.1080/19331681.2025.2473994.
- Cotra (2023). Language models surprised us.Cotra. (2023). Language models surprised us. https://www.planned-obsolescence.org/language-models-surprised-usCotra. 2023. “Language Models Surprised Us”. https://www.planned-obsolescence.org/language-models-surprised-us.Cotra. Language Models Surprised Us. 2023, https://www.planned-obsolescence.org/language-models-surprised-us.Cotra. Language models surprised us. https://www.planned-obsolescence.org/language-models-surprised-us (2023).Cotra, “Language models surprised us”. [Online]. Available: https://www.planned-obsolescence.org/language-models-surprised-us
- Cullen O’Keefe, Peter Cihon, Ben Garfinkel, Carrick Flynn, Jade Leung & and Allan Dafoe (2020). The Windfall Clause: Distributing the Benefits of AI for the Common Good.Cullen O’Keefe, Peter Cihon, Ben Garfinkel, Carrick Flynn, Jade Leung, & and Allan Dafoe. (2020, January 30). The Windfall Clause: Distributing the Benefits of AI for the Common Good. https://governance.ai/research-paper/the-windfall-clause-distributing-the-benefits-of-ai-for-the-common-goodCullen O’Keefe, Peter Cihon, Ben Garfinkel, Carrick Flynn, Jade Leung, and and Allan Dafoe. 2020. “The Windfall Clause: Distributing the Benefits of AI for the Common Good”. January 30. https://governance.ai/research-paper/the-windfall-clause-distributing-the-benefits-of-ai-for-the-common-good.Cullen O’Keefe, et al. The Windfall Clause: Distributing the Benefits of AI for the Common Good. 30 Jan. 2020, https://governance.ai/research-paper/the-windfall-clause-distributing-the-benefits-of-ai-for-the-common-good.Cullen O’Keefe et al. The Windfall Clause: Distributing the Benefits of AI for the Common Good. https://governance.ai/research-paper/the-windfall-clause-distributing-the-benefits-of-ai-for-the-common-good (2020).Cullen O’Keefe, Peter Cihon, Ben Garfinkel, Carrick Flynn, Jade Leung, and and Allan Dafoe, “The Windfall Clause: Distributing the Benefits of AI for the Common Good”. [Online]. Available: https://governance.ai/research-paper/the-windfall-clause-distributing-the-benefits-of-ai-for-the-common-good
- DeepSeek (2025). GitHub - deepseek-ai/DeepSeek-V3. GitHub.DeepSeek. (2025). GitHub - deepseek-ai/DeepSeek-V3. GitHub. https://github.com/deepseek-ai/DeepSeek-V3DeepSeek. 2025. “GitHub - deepseek-ai/DeepSeek-V3”. GitHub. https://github.com/deepseek-ai/DeepSeek-V3.DeepSeek. “GitHub - deepseek-ai/DeepSeek-V3”. GitHub, 2025, https://github.com/deepseek-ai/DeepSeek-V3.DeepSeek. GitHub - deepseek-ai/DeepSeek-V3. GitHub https://github.com/deepseek-ai/DeepSeek-V3 (2025).DeepSeek, “GitHub - deepseek-ai/DeepSeek-V3”, GitHub. [Online]. Available: https://github.com/deepseek-ai/DeepSeek-V3
- Ding, J. & Dafoe, A. (2020). The Logic of Strategic Assets: From Oil to Artificial Intelligence. arXiv.Ding, J., & Dafoe, A. (2020). The Logic of Strategic Assets: From Oil to Artificial Intelligence. In arXiv. https://doi.org/10.1080/09636412.2021.1915583Ding, J., and A. Dafoe. 2020. “The Logic of Strategic Assets: From Oil to Artificial Intelligence”. In arXiv. Preprint, January 9. https://doi.org/10.1080/09636412.2021.1915583.Ding, J., and A. Dafoe. “The Logic of Strategic Assets: From Oil to Artificial Intelligence”. arXiv, 9 Jan. 2020, https://doi.org/10.1080/09636412.2021.1915583.Ding, J. & Dafoe, A. The Logic of Strategic Assets: From Oil to Artificial Intelligence. arXiv Preprint at https://doi.org/10.1080/09636412.2021.1915583 (2020).J. Ding and A. Dafoe, “The Logic of Strategic Assets: From Oil to Artificial Intelligence”, Jan. 09, 2020. doi: 10.1080/09636412.2021.1915583.
- Dong, Y. et al. (2024). Safeguarding Large Language Models: A Survey. arXiv.Dong, Y., Mu, R., Zhang, Y., Sun, S., Zhang, T., Wu, C., Jin, G., Qi, Y., Hu, J., Meng, J., Bensalem, S., & Huang, X. (2024). Safeguarding Large Language Models: A Survey. In arXiv. https://arxiv.org/abs/2406.02622Dong, Y., R. Mu, Y. Zhang, et al. 2024. “Safeguarding Large Language Models: A Survey”. In arXiv. Preprint, June 3. https://arxiv.org/abs/2406.02622.Dong, Y., et al. “Safeguarding Large Language Models: A Survey”. arXiv, 3 June 2024, https://arxiv.org/abs/2406.02622.Dong, Y. et al. Safeguarding Large Language Models: A Survey. arXiv Preprint at https://arxiv.org/abs/2406.02622 (2024).Y. Dong et al., “Safeguarding Large Language Models: A Survey”, Jun. 03, 2024. [Online]. Available: https://arxiv.org/abs/2406.02622
- Eiras, F. et al. (2024). Near to Mid-term Risks and Opportunities of Open-Source Generative AI. arXiv.Eiras, F., Petrov, A., Vidgen, B., de Witt, C. S., Pizzati, F., Elkins, K., Mukhopadhyay, S., Bibi, A., Csaba, B., Steibel, F., Barez, F., Smith, G., Guadagni, G., Chun, J., Cabot, J., Imperial, J. M., Nolazco-Flores, J. A., Landay, L., Jackson, M., … Foerster, J. (2024). Near to Mid-term Risks and Opportunities of Open-Source Generative AI. In arXiv. https://arxiv.org/abs/2404.17047Eiras, F., A. Petrov, B. Vidgen, et al. 2024. “Near to Mid-term Risks and Opportunities of Open-Source Generative AI”. In arXiv. Preprint, April 25. https://arxiv.org/abs/2404.17047.Eiras, F., et al. “Near to Mid-term Risks and Opportunities of Open-Source Generative AI”. arXiv, 25 Apr. 2024, https://arxiv.org/abs/2404.17047.Eiras, F. et al. Near to Mid-term Risks and Opportunities of Open-Source Generative AI. arXiv Preprint at https://arxiv.org/abs/2404.17047 (2024).F. Eiras et al., “Near to Mid-term Risks and Opportunities of Open-Source Generative AI”, Apr. 25, 2024. [Online]. Available: https://arxiv.org/abs/2404.17047
- Giattino et al. (2023). Artificial Intelligence. Our World in Data.Giattino et al. (2023). Artificial Intelligence. Our World in Data. https://ourworldindata.org/artificial-intelligenceGiattino et al. 2023. “Artificial Intelligence”. Our World in Data. https://ourworldindata.org/artificial-intelligence.Giattino et al. “Artificial Intelligence”. Our World in Data, 2023, https://ourworldindata.org/artificial-intelligence.Giattino et al. Artificial Intelligence. Our World in Data https://ourworldindata.org/artificial-intelligence (2023).Giattino et al., “Artificial Intelligence”, Our World in Data. [Online]. Available: https://ourworldindata.org/artificial-intelligence
- Grace, K. et al. (2024). Thousands of AI Authors on the Future of AI. arXiv.Grace, K., Stewart, H., Sandkühler, J. F., Thomas, S., Weinstein-Raun, B., Brauner, J., & Korzekwa, R. C. (2024). Thousands of AI Authors on the Future of AI. In arXiv. https://doi.org/10.1613/jair.1.19087Grace, K., H. Stewart, J. F. Sandkühler, et al. 2024. “Thousands of AI Authors on the Future of AI”. In arXiv. Preprint, January 5. https://doi.org/10.1613/jair.1.19087.Grace, K., et al. “Thousands of AI Authors on the Future of AI”. arXiv, 5 Jan. 2024, https://doi.org/10.1613/jair.1.19087.Grace, K. et al. Thousands of AI Authors on the Future of AI. arXiv Preprint at https://doi.org/10.1613/jair.1.19087 (2024).K. Grace et al., “Thousands of AI Authors on the Future of AI”, Jan. 05, 2024. doi: 10.1613/jair.1.19087.
- Grace, K., Salvatier, J., Dafoe, A., Zhang, B. & Evans, O. (2017). When Will AI Exceed Human Performance? Evidence from AI Experts. arXiv.Grace, K., Salvatier, J., Dafoe, A., Zhang, B., & Evans, O. (2017). When Will AI Exceed Human Performance? Evidence from AI Experts. In arXiv. https://arxiv.org/abs/1705.08807Grace, K., J. Salvatier, A. Dafoe, B. Zhang, and O. Evans. 2017. “When Will AI Exceed Human Performance? Evidence from AI Experts”. In arXiv. Preprint, May 24. https://arxiv.org/abs/1705.08807.Grace, K., et al. “When Will AI Exceed Human Performance? Evidence from AI Experts”. arXiv, 24 May 2017, https://arxiv.org/abs/1705.08807.Grace, K., Salvatier, J., Dafoe, A., Zhang, B. & Evans, O. When Will AI Exceed Human Performance? Evidence from AI Experts. arXiv Preprint at https://arxiv.org/abs/1705.08807 (2017).K. Grace, J. Salvatier, A. Dafoe, B. Zhang, and O. Evans, “When Will AI Exceed Human Performance? Evidence from AI Experts”, May 24, 2017. [Online]. Available: https://arxiv.org/abs/1705.08807
- Gruetzemacher et al. (2021). Forecasting AI progress: A research agenda.Gruetzemacher et al. (2021). Forecasting AI progress: A research agenda. Internet Archive (https://web.archive.org/web/20240404122334/https://www.sciencedirect.com/science/article/pii/S0040162521003413). https://sciencedirect.com/science/article/pii/S0040162521003413Gruetzemacher et al. 2021. “Forecasting AI Progress: A Research Agenda”. Https://web.archive.org/web/20240404122334/https://www.sciencedirect.com/science/article/pii/S0040162521003413. Internet Archive. https://sciencedirect.com/science/article/pii/S0040162521003413.Gruetzemacher et al. Forecasting AI Progress: A Research Agenda. 2021, Internet Archive, https://web.archive.org/web/20240404122334/https://www.sciencedirect.com/science/article/pii/S0040162521003413, https://sciencedirect.com/science/article/pii/S0040162521003413.Gruetzemacher et al. Forecasting AI progress: A research agenda. https://sciencedirect.com/science/article/pii/S0040162521003413 (2021).Gruetzemacher et al., “Forecasting AI progress: A research agenda”. Accessed: Apr. 04, 2024. [Online]. Available: https://sciencedirect.com/science/article/pii/S0040162521003413
- Gruetzemacher, R., Avin, S., Fox, J. & Saeri, A. K. (2024). Strategic Insights from Simulation Gaming of AI Race Dynamics. arXiv.Gruetzemacher, R., Avin, S., Fox, J., & Saeri, A. K. (2024). Strategic Insights from Simulation Gaming of AI Race Dynamics. In arXiv. https://arxiv.org/abs/2410.03092Gruetzemacher, R., S. Avin, J. Fox, and A. K. Saeri. 2024. “Strategic Insights from Simulation Gaming of AI Race Dynamics”. In arXiv. Preprint, October 4. https://arxiv.org/abs/2410.03092.Gruetzemacher, R., et al. “Strategic Insights from Simulation Gaming of AI Race Dynamics”. arXiv, 4 Oct. 2024, https://arxiv.org/abs/2410.03092.Gruetzemacher, R., Avin, S., Fox, J. & Saeri, A. K. Strategic Insights from Simulation Gaming of AI Race Dynamics. arXiv Preprint at https://arxiv.org/abs/2410.03092 (2024).R. Gruetzemacher, S. Avin, J. Fox, and A. K. Saeri, “Strategic Insights from Simulation Gaming of AI Race Dynamics”, Oct. 04, 2024. [Online]. Available: https://arxiv.org/abs/2410.03092
- Hendryks (2024). 7.2: Game Theory | AI Safety, Ethics, and Society Textbook.Hendryks. (2024). 7.2: Game Theory | AI Safety, Ethics, and Society Textbook. https://aisafetybook.com/textbook/game-theoryHendryks. 2024. “7.2: Game Theory | AI Safety, Ethics, and Society Textbook”. https://aisafetybook.com/textbook/game-theory.Hendryks. 7.2: Game Theory | AI Safety, Ethics, and Society Textbook. 2024, https://aisafetybook.com/textbook/game-theory.Hendryks. 7.2: Game Theory | AI Safety, Ethics, and Society Textbook. https://aisafetybook.com/textbook/game-theory (2024).Hendryks, “7.2: Game Theory | AI Safety, Ethics, and Society Textbook”. [Online]. Available: https://aisafetybook.com/textbook/game-theory
- Herrera-Poyatos, A., Ser, J. D., Prado, M. L. D., Wang, F., Herrera-Viedma, E. & Herrera, F. (2025). A Framework for Responsible AI Systems: Building Societal Trust through Domain Definition, Trustworthy AI Design, Auditability, Accountability, and Governance. arXiv.Herrera-Poyatos, A., Ser, J. D., de Prado, M. L., Wang, F.-Y., Herrera-Viedma, E., & Herrera, F. (2025). A Framework for Responsible AI Systems: Building Societal Trust through Domain Definition, Trustworthy AI Design, Auditability, Accountability, and Governance. In arXiv. https://arxiv.org/abs/2503.04739Herrera-Poyatos, A., J. D. Ser, M. L. de Prado, F.-Y. Wang, E. Herrera-Viedma, and F. Herrera. 2025. “A Framework for Responsible AI Systems: Building Societal Trust Through Domain Definition, Trustworthy AI Design, Auditability, Accountability, and Governance”. In arXiv. Preprint, February 4. https://arxiv.org/abs/2503.04739.Herrera-Poyatos, A., et al. “A Framework for Responsible AI Systems: Building Societal Trust Through Domain Definition, Trustworthy AI Design, Auditability, Accountability, and Governance”. arXiv, 4 Feb. 2025, https://arxiv.org/abs/2503.04739.Herrera-Poyatos, A. et al. A Framework for Responsible AI Systems: Building Societal Trust through Domain Definition, Trustworthy AI Design, Auditability, Accountability, and Governance. arXiv Preprint at https://arxiv.org/abs/2503.04739 (2025).A. Herrera-Poyatos, J. D. Ser, M. L. de Prado, F.-Y. Wang, E. Herrera-Viedma, and F. Herrera, “A Framework for Responsible AI Systems: Building Societal Trust through Domain Definition, Trustworthy AI Design, Auditability, Accountability, and Governance”, Feb. 04, 2025. [Online]. Available: https://arxiv.org/abs/2503.04739
- Joe O'Brien (2024). Coordinated Disclosure of Dual-Use Capabilities: An Early Warning System for Advanced AI.Joe O'Brien. (2024, June 21). Coordinated Disclosure of Dual-Use Capabilities: An Early Warning System for Advanced AI. https://iaps.ai/research/coordinated-disclosureJoe O'Brien. 2024. “Coordinated Disclosure of Dual-Use Capabilities: An Early Warning System for Advanced AI”. June 21. https://iaps.ai/research/coordinated-disclosure.Joe O'Brien. Coordinated Disclosure of Dual-Use Capabilities: An Early Warning System for Advanced AI. 21 June 2024, https://iaps.ai/research/coordinated-disclosure.Joe O'Brien. Coordinated Disclosure of Dual-Use Capabilities: An Early Warning System for Advanced AI. https://iaps.ai/research/coordinated-disclosure (2024).Joe O'Brien, “Coordinated Disclosure of Dual-Use Capabilities: An Early Warning System for Advanced AI”. [Online]. Available: https://iaps.ai/research/coordinated-disclosure
- Karnofsky (2024). If-Then Commitments for AI Risk Reduction. Carnegie Endowment for International Peace.Karnofsky. (2024, September 13). If-Then Commitments for AI Risk Reduction. Carnegie Endowment for International Peace. https://carnegieendowment.org/research/2024/09/if-then-commitments-for-ai-risk-reduction?lang=enKarnofsky. 2024. “If-Then Commitments for AI Risk Reduction”. Carnegie Endowment for International Peace, September 13. https://carnegieendowment.org/research/2024/09/if-then-commitments-for-ai-risk-reduction?lang=en.Karnofsky. “If-Then Commitments for AI Risk Reduction”. Carnegie Endowment for International Peace, 13 Sept. 2024, https://carnegieendowment.org/research/2024/09/if-then-commitments-for-ai-risk-reduction?lang=en.Karnofsky. If-Then Commitments for AI Risk Reduction. Carnegie Endowment for International Peace https://carnegieendowment.org/research/2024/09/if-then-commitments-for-ai-risk-reduction?lang=en (2024).Karnofsky, “If-Then Commitments for AI Risk Reduction”, Carnegie Endowment for International Peace. [Online]. Available: https://carnegieendowment.org/research/2024/09/if-then-commitments-for-ai-risk-reduction?lang=en
- Kreps & Kriner (2023). How AI Threatens Democracy. Journal of Democracy.Kreps & Kriner. (2023). How AI Threatens Democracy. Journal of Democracy. https://journalofdemocracy.org/articles/how-ai-threatens-democracyKreps & Kriner. 2023. “How AI Threatens Democracy”. Journal of Democracy. https://journalofdemocracy.org/articles/how-ai-threatens-democracy.Kreps & Kriner. “How AI Threatens Democracy”. Journal of Democracy, 2023, https://journalofdemocracy.org/articles/how-ai-threatens-democracy.Kreps & Kriner. How AI Threatens Democracy. Journal of Democracy https://journalofdemocracy.org/articles/how-ai-threatens-democracy (2023).Kreps & Kriner, “How AI Threatens Democracy”, Journal of Democracy. [Online]. Available: https://journalofdemocracy.org/articles/how-ai-threatens-democracy
- Lazar, S. (2024). Automatic Authorities: Power and AI. arXiv.Lazar, S. (2024). Automatic Authorities: Power and AI. In arXiv. https://arxiv.org/abs/2404.05990Lazar, S. 2024. “Automatic Authorities: Power and AI”. In arXiv. Preprint, April 9. https://arxiv.org/abs/2404.05990.Lazar, S. “Automatic Authorities: Power and AI”. arXiv, 9 Apr. 2024, https://arxiv.org/abs/2404.05990.Lazar, S. Automatic Authorities: Power and AI. arXiv Preprint at https://arxiv.org/abs/2404.05990 (2024).S. Lazar, “Automatic Authorities: Power and AI”, Apr. 09, 2024. [Online]. Available: https://arxiv.org/abs/2404.05990
- Masi (2024). Masi_GPU_Export_Controls_Updated.pdf. Google Docs.Masi. (2024). Masi_GPU_Export_Controls_Updated.pdf. Google Docs. https://drive.google.com/file/d/1SSRpckR_IhIPqkO13oRQ-8mmPOoDKLqd/viewMasi. 2024. “Masi_GPU_Export_Controls_Updated.pdf”. Google Docs. https://drive.google.com/file/d/1SSRpckR_IhIPqkO13oRQ-8mmPOoDKLqd/view.Masi. “Masi_GPU_Export_Controls_Updated.pdf”. Google Docs, 2024, https://drive.google.com/file/d/1SSRpckR_IhIPqkO13oRQ-8mmPOoDKLqd/view.Masi. Masi_GPU_Export_Controls_Updated.pdf. Google Docs https://drive.google.com/file/d/1SSRpckR_IhIPqkO13oRQ-8mmPOoDKLqd/view (2024).Masi, “Masi_GPU_Export_Controls_Updated.pdf”, Google Docs. [Online]. Available: https://drive.google.com/file/d/1SSRpckR_IhIPqkO13oRQ-8mmPOoDKLqd/view
- Nevo et al. (2024). How AI Labs Can Safeguard Model Weights.Nevo et al. (2024). How AI Labs Can Safeguard Model Weights. Internet Archive (https://web.archive.org/web/20260920224534/https://www.rand.org/pubs/research_reports/RRA2849-1.html). https://rand.org/pubs/research_reports/RRA2849-1.htmlNevo et al. 2024. “How AI Labs Can Safeguard Model Weights”. Https://web.archive.org/web/20260920224534/https://www.rand.org/pubs/research_reports/RRA2849-1.html. Internet Archive. https://rand.org/pubs/research_reports/RRA2849-1.html.Nevo et al. How AI Labs Can Safeguard Model Weights. 2024, Internet Archive, https://web.archive.org/web/20260920224534/https://www.rand.org/pubs/research_reports/RRA2849-1.html, https://rand.org/pubs/research_reports/RRA2849-1.html.Nevo et al. How AI Labs Can Safeguard Model Weights. https://rand.org/pubs/research_reports/RRA2849-1.html (2024).Nevo et al., “How AI Labs Can Safeguard Model Weights”. Accessed: Sep. 20, 2026. [Online]. Available: https://rand.org/pubs/research_reports/RRA2849-1.html
- Patel (2023). Will scaling work?.Patel. (2023). Will scaling work?. https://dwarkesh.com/p/will-scaling-workPatel. 2023. “Will Scaling Work?”. https://dwarkesh.com/p/will-scaling-work.Patel. Will Scaling Work?. 2023, https://dwarkesh.com/p/will-scaling-work.Patel. Will scaling work?. https://dwarkesh.com/p/will-scaling-work (2023).Patel, “Will scaling work?”. [Online]. Available: https://dwarkesh.com/p/will-scaling-work
- Petropoulos et al. (2025). Building CERN for AI - An institutional blueprint. Centre for Future Generations.Petropoulos et al. (2025, January 30). Building CERN for AI - An institutional blueprint. Centre for Future Generations. https://cfg.eu/building-cern-for-aiPetropoulos et al. 2025. “Building CERN for AI - An Institutional Blueprint”. Centre for Future Generations, January 30. https://cfg.eu/building-cern-for-ai.Petropoulos et al. “Building CERN for AI - An Institutional Blueprint”. Centre for Future Generations, 30 Jan. 2025, https://cfg.eu/building-cern-for-ai.Petropoulos et al. Building CERN for AI - An institutional blueprint. Centre for Future Generations https://cfg.eu/building-cern-for-ai (2025).Petropoulos et al., “Building CERN for AI - An institutional blueprint”, Centre for Future Generations. [Online]. Available: https://cfg.eu/building-cern-for-ai
- Rishub Tamirisa et al. (2024). Tamper-Resistant Safeguards for Open-Weight LLMs. arXiv.Rishub Tamirisa, Bhrugu Bharathi, Long Phan, Andy Zhou, Alice Gatti, Tarun Suresh, Maxwell Lin, Justin Wang, Rowan Wang, Ron Arel, Andy Zou, Dawn Song, Bo Li, Dan Hendrycks, & Mantas Mazeika. (2024). Tamper-Resistant Safeguards for Open-Weight LLMs. In arXiv. https://arxiv.org/abs/2408.00761Rishub Tamirisa, Bhrugu Bharathi, Long Phan, et al. 2024. “Tamper-Resistant Safeguards for Open-Weight LLMs”. In arXiv. Preprint, August 1. https://arxiv.org/abs/2408.00761.Rishub Tamirisa, et al. “Tamper-Resistant Safeguards for Open-Weight LLMs”. arXiv, 1 Aug. 2024, https://arxiv.org/abs/2408.00761.Rishub Tamirisa et al. Tamper-Resistant Safeguards for Open-Weight LLMs. arXiv Preprint at https://arxiv.org/abs/2408.00761 (2024).Rishub Tamirisa et al., “Tamper-Resistant Safeguards for Open-Weight LLMs”, Aug. 01, 2024. [Online]. Available: https://arxiv.org/abs/2408.00761
- Seger, E. et al. (2023). Open-Sourcing Highly Capable Foundation Models: An evaluation of risks, benefits, and alternative methods for pursuing open-source objectives. arXiv.Seger, E., Dreksler, N., Moulange, R., Dardaman, E., Schuett, J., Wei, K., Winter, C., Arnold, M., hÉigeartaigh, S. Ó., Korinek, A., Anderljung, M., Bucknall, B., Chan, A., Stafford, E., Koessler, L., Ovadya, A., Garfinkel, B., Bluemke, E., Aird, M., … Gupta, A. (2023). Open-Sourcing Highly Capable Foundation Models: An evaluation of risks, benefits, and alternative methods for pursuing open-source objectives. In arXiv. https://arxiv.org/abs/2311.09227Seger, E., N. Dreksler, R. Moulange, et al. 2023. “Open-Sourcing Highly Capable Foundation Models: An Evaluation of Risks, Benefits, and Alternative Methods for Pursuing Open-source Objectives”. In arXiv. Preprint, September 29. https://arxiv.org/abs/2311.09227.Seger, E., et al. “Open-Sourcing Highly Capable Foundation Models: An Evaluation of Risks, Benefits, and Alternative Methods for Pursuing Open-source Objectives”. arXiv, 29 Sept. 2023, https://arxiv.org/abs/2311.09227.Seger, E. et al. Open-Sourcing Highly Capable Foundation Models: An evaluation of risks, benefits, and alternative methods for pursuing open-source objectives. arXiv Preprint at https://arxiv.org/abs/2311.09227 (2023).E. Seger et al., “Open-Sourcing Highly Capable Foundation Models: An evaluation of risks, benefits, and alternative methods for pursuing open-source objectives”, Sep. 29, 2023. [Online]. Available: https://arxiv.org/abs/2311.09227
- Shevlane, T. & Dafoe, A. (2019). The Offense-Defense Balance of Scientific Knowledge: Does Publishing AI Research Reduce Misuse?. arXiv.Shevlane, T., & Dafoe, A. (2019). The Offense-Defense Balance of Scientific Knowledge: Does Publishing AI Research Reduce Misuse?. In arXiv. https://arxiv.org/abs/2001.00463Shevlane, T., and A. Dafoe. 2019. “The Offense-Defense Balance of Scientific Knowledge: Does Publishing AI Research Reduce Misuse?”. In arXiv. Preprint, December 27. https://arxiv.org/abs/2001.00463.Shevlane, T., and A. Dafoe. “The Offense-Defense Balance of Scientific Knowledge: Does Publishing AI Research Reduce Misuse?”. arXiv, 27 Dec. 2019, https://arxiv.org/abs/2001.00463.Shevlane, T. & Dafoe, A. The Offense-Defense Balance of Scientific Knowledge: Does Publishing AI Research Reduce Misuse?. arXiv Preprint at https://arxiv.org/abs/2001.00463 (2019).T. Shevlane and A. Dafoe, “The Offense-Defense Balance of Scientific Knowledge: Does Publishing AI Research Reduce Misuse?”, Dec. 27, 2019. [Online]. Available: https://arxiv.org/abs/2001.00463
- Skare et al. (2024). Artificial intelligence and wealth inequality: A comprehensi.Skare et al. (2024). Artificial intelligence and wealth inequality: A comprehensi. https://ideas.repec.org/a/eee/teinso/v79y2024ics0160791x24002677.htmlSkare et al. 2024. “Artificial Intelligence and Wealth Inequality: A Comprehensi”. https://ideas.repec.org/a/eee/teinso/v79y2024ics0160791x24002677.html.Skare et al. Artificial Intelligence and Wealth Inequality: A Comprehensi. 2024, https://ideas.repec.org/a/eee/teinso/v79y2024ics0160791x24002677.html.Skare et al. Artificial intelligence and wealth inequality: A comprehensi. https://ideas.repec.org/a/eee/teinso/v79y2024ics0160791x24002677.html (2024).Skare et al., “Artificial intelligence and wealth inequality: A comprehensi”. [Online]. Available: https://ideas.repec.org/a/eee/teinso/v79y2024ics0160791x24002677.html
- Solaiman, I. (2023). The Gradient of Generative AI Release: Methods and Considerations. arXiv.Solaiman, I. (2023). The Gradient of Generative AI Release: Methods and Considerations. In arXiv. https://arxiv.org/abs/2302.04844Solaiman, I. 2023. “The Gradient of Generative AI Release: Methods and Considerations”. In arXiv. Preprint, February 5. https://arxiv.org/abs/2302.04844.Solaiman, I. “The Gradient of Generative AI Release: Methods and Considerations”. arXiv, 5 Feb. 2023, https://arxiv.org/abs/2302.04844.Solaiman, I. The Gradient of Generative AI Release: Methods and Considerations. arXiv Preprint at https://arxiv.org/abs/2302.04844 (2023).I. Solaiman, “The Gradient of Generative AI Release: Methods and Considerations”, Feb. 05, 2023. [Online]. Available: https://arxiv.org/abs/2302.04844
- Stacey (2025). AI-generated phishing scams target corporate executives.Stacey. (2025). AI-generated phishing scams target corporate executives. https://ft.com/content/d60fb4fb-cb85-4df7-b246-ec3d08260e6fStacey. 2025. “AI-generated Phishing Scams Target Corporate Executives”. https://ft.com/content/d60fb4fb-cb85-4df7-b246-ec3d08260e6f.Stacey. AI-generated Phishing Scams Target Corporate Executives. 2025, https://ft.com/content/d60fb4fb-cb85-4df7-b246-ec3d08260e6f.Stacey. AI-generated phishing scams target corporate executives. https://ft.com/content/d60fb4fb-cb85-4df7-b246-ec3d08260e6f (2025).Stacey, “AI-generated phishing scams target corporate executives”. [Online]. Available: https://ft.com/content/d60fb4fb-cb85-4df7-b246-ec3d08260e6f
- Stanford (2024). The 2024 AI Index Report. Stanford HAI.Stanford. (2024). The 2024 AI Index Report. Stanford HAI. https://hai.stanford.edu/ai-index/2024-ai-index-reportStanford. 2024. “The 2024 AI Index Report”. Stanford HAI. https://hai.stanford.edu/ai-index/2024-ai-index-report.Stanford. “The 2024 AI Index Report”. Stanford HAI, 2024, https://hai.stanford.edu/ai-index/2024-ai-index-report.Stanford. The 2024 AI Index Report. Stanford HAI https://hai.stanford.edu/ai-index/2024-ai-index-report (2024).Stanford, “The 2024 AI Index Report”, Stanford HAI. [Online]. Available: https://hai.stanford.edu/ai-index/2024-ai-index-report
- Stanford HAI (2024). AI Index | Stanford HAI.Stanford HAI. (2024). AI Index | Stanford HAI. https://aiindex.stanford.edu/reportStanford HAI. 2024. “AI Index | Stanford HAI”. https://aiindex.stanford.edu/report.Stanford HAI. AI Index | Stanford HAI. 2024, https://aiindex.stanford.edu/report.Stanford HAI. AI Index | Stanford HAI. https://aiindex.stanford.edu/report (2024).Stanford HAI, “AI Index | Stanford HAI”. [Online]. Available: https://aiindex.stanford.edu/report
- Stanford HAI (2025). The 2025 AI Index Report. Stanford HAI.Stanford HAI. (2025). The 2025 AI Index Report. Stanford HAI. https://hai.stanford.edu/ai-index/2025-ai-index-reportStanford HAI. 2025. “The 2025 AI Index Report”. Stanford HAI. https://hai.stanford.edu/ai-index/2025-ai-index-report.Stanford HAI. “The 2025 AI Index Report”. Stanford HAI, 2025, https://hai.stanford.edu/ai-index/2025-ai-index-report.Stanford HAI. The 2025 AI Index Report. Stanford HAI https://hai.stanford.edu/ai-index/2025-ai-index-report (2025).Stanford HAI, “The 2025 AI Index Report”, Stanford HAI. [Online]. Available: https://hai.stanford.edu/ai-index/2025-ai-index-report
- Stanford Institute for Human-Centered Artificial Intelligence (2025). The AI Index 2025 Annual Report.Stanford Institute for Human-Centered Artificial Intelligence. (2025). The AI Index 2025 Annual Report. Stanford Institute for Human-Centered Artificial Intelligence. https://doi.org/10.48550/arXiv.2504.07139Stanford Institute for Human-Centered Artificial Intelligence. 2025. The AI Index 2025 Annual Report. Stanford Institute for Human-Centered Artificial Intelligence. https://doi.org/10.48550/arXiv.2504.07139.Stanford Institute for Human-Centered Artificial Intelligence. The AI Index 2025 Annual Report. Stanford Institute for Human-Centered Artificial Intelligence, Apr. 2025, https://doi.org/10.48550/arXiv.2504.07139.Stanford Institute for Human-Centered Artificial Intelligence. The AI Index 2025 Annual Report. https://hai.stanford.edu/assets/files/hai_ai_index_report_2025.pdf (2025) doi:10.48550/arXiv.2504.07139.Stanford Institute for Human-Centered Artificial Intelligence, “The AI Index 2025 Annual Report”, Stanford Institute for Human-Centered Artificial Intelligence, Stanford, CA, Apr. 2025. doi: 10.48550/arXiv.2504.07139.
- Stewart, A. J. & Plotkin, J. B. (2012). Extortion and cooperation in the Prisoner’s Dilemma. Proceedings of the National Academy of Sciences.Stewart, A. J., & Plotkin, J. B. (2012). Extortion and cooperation in the Prisoner’s Dilemma. Proceedings of the National Academy of Sciences. https://doi.org/10.1073/pnas.1208087109Stewart, A. J., and J. B. Plotkin. 2012. “Extortion and Cooperation in the Prisoner’s Dilemma”. Proceedings of the National Academy of Sciences, ahead of print, June 18. https://doi.org/10.1073/pnas.1208087109.Stewart, A. J., and J. B. Plotkin. “Extortion and Cooperation in the Prisoner’s Dilemma”. Proceedings of the National Academy of Sciences, June 2012, https://doi.org/10.1073/pnas.1208087109.Stewart, A. J. & Plotkin, J. B. Extortion and cooperation in the Prisoner’s Dilemma. Proceedings of the National Academy of Sciences https://doi.org/10.1073/pnas.1208087109 (2012) doi:10.1073/pnas.1208087109.A. J. Stewart and J. B. Plotkin, “Extortion and cooperation in the Prisoner’s Dilemma”, Proceedings of the National Academy of Sciences, Jun. 2012, doi: 10.1073/pnas.1208087109.
- Stix, C. et al. (2025). AI Behind Closed Doors: a Primer on The Governance of Internal Deployment. arXiv.Stix, C., Pistillo, M., Sastry, G., Hobbhahn, M., Ortega, A., Balesni, M., Hallensleben, A., Goldowsky-Dill, N., & Sharkey, L. (2025). AI Behind Closed Doors: a Primer on The Governance of Internal Deployment. In arXiv. https://arxiv.org/abs/2504.12170Stix, C., M. Pistillo, G. Sastry, et al. 2025. “AI Behind Closed Doors: A Primer on The Governance of Internal Deployment”. In arXiv. Preprint, April 16. https://arxiv.org/abs/2504.12170.Stix, C., et al. “AI Behind Closed Doors: A Primer on The Governance of Internal Deployment”. arXiv, 16 Apr. 2025, https://arxiv.org/abs/2504.12170.Stix, C. et al. AI Behind Closed Doors: a Primer on The Governance of Internal Deployment. arXiv Preprint at https://arxiv.org/abs/2504.12170 (2025).C. Stix et al., “AI Behind Closed Doors: a Primer on The Governance of Internal Deployment”, Apr. 16, 2025. [Online]. Available: https://arxiv.org/abs/2504.12170
- Swenson & Chan (2024). Election disinformation takes a big leap with AI being used to deceive worldwide. AP News.Swenson & Chan. (2024, March 14). Election disinformation takes a big leap with AI being used to deceive worldwide. AP News. https://apnews.com/article/artificial-intelligence-elections-disinformation-chatgpt-bc283e7426402f0b4baa7df280a4c3fdSwenson & Chan. 2024. “Election Disinformation Takes a Big Leap with AI Being Used to Deceive Worldwide”. AP News, March 14. https://apnews.com/article/artificial-intelligence-elections-disinformation-chatgpt-bc283e7426402f0b4baa7df280a4c3fd.Swenson & Chan. “Election Disinformation Takes a Big Leap with AI Being Used to Deceive Worldwide”. AP News, 14 Mar. 2024, https://apnews.com/article/artificial-intelligence-elections-disinformation-chatgpt-bc283e7426402f0b4baa7df280a4c3fd.Swenson & Chan. Election disinformation takes a big leap with AI being used to deceive worldwide. AP News https://apnews.com/article/artificial-intelligence-elections-disinformation-chatgpt-bc283e7426402f0b4baa7df280a4c3fd (2024).Swenson & Chan, “Election disinformation takes a big leap with AI being used to deceive worldwide”, AP News. [Online]. Available: https://apnews.com/article/artificial-intelligence-elections-disinformation-chatgpt-bc283e7426402f0b4baa7df280a4c3fd
- Trajano & Ang (2023). We Need to Prevent a Global AI Arms Race Now. RSIS_NTU.Trajano & Ang. (2023). We Need to Prevent a Global AI Arms Race Now. RSIS_NTU. https://rsis.edu.sg/rsis-publication/rsis/we-need-to-prevent-a-global-ai-arms-race-nowTrajano & Ang. 2023. “We Need to Prevent a Global AI Arms Race Now”. RSIS_NTU. https://rsis.edu.sg/rsis-publication/rsis/we-need-to-prevent-a-global-ai-arms-race-now.Trajano & Ang. “We Need to Prevent a Global AI Arms Race Now”. RSIS_NTU, 2023, https://rsis.edu.sg/rsis-publication/rsis/we-need-to-prevent-a-global-ai-arms-race-now.Trajano & Ang. We Need to Prevent a Global AI Arms Race Now. RSIS_NTU https://rsis.edu.sg/rsis-publication/rsis/we-need-to-prevent-a-global-ai-arms-race-now (2023).Trajano & Ang, “We Need to Prevent a Global AI Arms Race Now”, RSIS_NTU. [Online]. Available: https://rsis.edu.sg/rsis-publication/rsis/we-need-to-prevent-a-global-ai-arms-race-now
- Wasil, A. R., Reed, T., Miller, J. W. & Barnett, P. (2024). Verification methods for international AI agreements. arXiv.Wasil, A. R., Reed, T., Miller, J. W., & Barnett, P. (2024). Verification methods for international AI agreements. In arXiv. https://arxiv.org/abs/2408.16074Wasil, A. R., T. Reed, J. W. Miller, and P. Barnett. 2024. “Verification Methods for International AI Agreements”. In arXiv. Preprint, August 28. https://arxiv.org/abs/2408.16074.Wasil, A. R., et al. “Verification Methods for International AI Agreements”. arXiv, 28 Aug. 2024, https://arxiv.org/abs/2408.16074.Wasil, A. R., Reed, T., Miller, J. W. & Barnett, P. Verification methods for international AI agreements. arXiv Preprint at https://arxiv.org/abs/2408.16074 (2024).A. R. Wasil, T. Reed, J. W. Miller, and P. Barnett, “Verification methods for international AI agreements”, Aug. 28, 2024. [Online]. Available: https://arxiv.org/abs/2408.16074
Was this section useful?
Thank you for your feedback
Your input helps improve the Atlas.