MB
Miles Brundage
cs.CYcs.AIcs.LGcs.CLcs.CVcs.CRcs.SDcs.SEeess.ASphysics.pop-ph
On Valency
published · living versionsW_y3274kcu·v1 · currentpublished
The Malicious Use of Artificial Intelligence: Forecasting, Prevention, and Mitigation
with Shahar Avin, Jack Clark, Helen Toner, Peter Eckersley +21
1 version
Preprints & journals
26 papers in the corpus · 2015–2026Underwriting the Agent Economy: The Blueprint for an AI Insurance Stack2607.11999v2 · Cristian Trout, Sanmi Koyejo, Sasha Romanosky et al.2026 · 0 citationsarXiv
Frontier AI Auditing: Toward Rigorous Third-Party Assessment of Safety and Security Practices at Leading AI Companies2601.11699v4 · Miles Brundage, Noemi Dreksler, Aidan Homewood et al.2026 · 1 citationarXiv
Machine Unlearning Doesn't Do What You Think: Lessons for Generative AI Policy and Research2412.06966v2 · A. Feder Cooper, Christopher A. Choquette-Choo, Miranda Bogen et al.2024 · 3 citationsarXiv
Verifying International Agreements on AI: Six Layers of Verification for Rules on Large-Scale AI Development and Deployment2507.15916v2 · Mauricio Baker, Gabriel Kulp, Oliver Marks et al.2025 · 0 citationsarXiv
Third-party compliance reviews for frontier AI safety frameworks2505.01643v2 · Aidan Homewood, Sophie Williams, Noemi Dreksler et al.2025 · 0 citationsarXiv
The Malicious Use of Artificial Intelligence: Forecasting, Prevention, and Mitigation1802.07228v2 · Miles Brundage, Shahar Avin, Jack Clark et al.2018 · 495 citationsarXivon Valency
GPT-4o System Card2410.21276v1 · OpenAI: Aaron Hurst, Adam Lerer, Adam P. Goucher et al.2024 · 165 citationsarXiv
Computing Power and the Governance of Artificial Intelligence2402.08797v1 · Girish Sastry, Lennart Heim, Haydn Belfield et al.2024 · 22 citationsarXiv
Report of the 1st Workshop on Generative AI and Law2311.06477v3 · A. Feder Cooper, Katherine Lee, James Grimmelmann et al.2023 · 6 citationsarXiv
Frontier AI Regulation: Managing Emerging Risks to Public Safety2307.03718v4 · Markus Anderljung, Joslyn Barnhart, Anton Korinek et al.2023 · 76 citationsarXiv
Confidence-Building Measures for Artificial Intelligence: Workshop Proceedings2308.00862v2 · Sarah Shoker, Andrew Reddie, Sarah Barrington et al.2023 · 13 citationsarXiv
International Institutions for Advanced AI2307.04699v2 · Lewis Ho, Joslyn Barnhart, Robert Trager et al.2023 · 19 citationsarXiv
A Hazard Analysis Framework for Code Synthesis Large Language Models2207.14157v1 · Heidy Khlaaf, Pamela Mishkin, Joshua Achiam et al.2022 · 9 citationsarXiv
Between Progress and Potential Impact of AI: the Neglected Dimensions1806.00610v2 · Fernando Mart'inez-Plumed, Shahar Avin, Miles Brundage et al.2018 · 1 citationarXiv
Filling gaps in trustworthy development of AI2112.07773v1 · Shahar Avin, Haydn Belfield, Miles Brundage et al.2021 · 55 citationsScience (2021) Vol 374, Issue 6573, pp. 1327-1329
Evaluating CLIP: Towards Characterization of Broader Capabilities and Downstream Implications2108.02818v1 · Sandhini Agarwal, Gretchen Krueger, Jack Clark et al.2021 · 36 citationsarXiv
Evaluating Large Language Models Trained on Code2107.03374v2 · Mark Chen, Jerry Tworek, Heewoo Jun et al.2021 · 1,468 citationsarXivon Valency
Understanding the Capabilities, Limitations, and Societal Impact of Large Language Models2102.02503v1 · Alex Tamkin, Miles Brundage, Jack Clark et al.2021 · 129 citationsarXiv
Toward Trustworthy AI Development: Mechanisms for Supporting Verifiable Claims2004.07213v2 · Miles Brundage, Shahar Avin, Jasmine Wang et al.2020 · 301 citationsarXiv
Release Strategies and the Social Impacts of Language Models1908.09203v2 · Irene Solaiman, Miles Brundage, Jack Clark et al.2019 · 282 citationsarXiv
The Role of Cooperation in Responsible AI Development1907.04534v1 · Amanda Askell, Miles Brundage, Gillian Hadfield2019 · 46 citationsarXiv
A Brief Survey of Deep Reinforcement Learning1708.05866v2 · Kai Arulkumaran, Marc Peter Deisenroth, Miles Brundage et al.2017 · 4,809 citationsarXiv
On the Impossibility of Supersized Machines1703.10987v1 · Ben Garfinkel, Miles Brundage, Daniel Filan et al.2017 · 1 citationarXiv
Smart Policies for Artificial Intelligence1608.08196v1 · Miles Brundage, Joanna Bryson2016 · 20 citationsarXivon Valency
Modeling Progress in AI1512.05849v1 · Miles Brundage2015 · 2 citationsarXiv
Career total: 41 works. 26 are in this corpus.Showing the 25 most recent.
Profile built from the corpus for this byline.
Author records are still filling in while the Hub is in alpha. If this is your page, you'll be able to claim it soon. Spot a mistake? Tell us.