AH
Ariel Herbert-Voss
cs.CLcs.CYcs.AIcs.LGcs.CR
On Valency
published · living versionsW_3seyyh2p·v1 · currentpublished
Evaluating Large Language Models Trained on Code
with Mark Chen, Jerry Tworek, Heewoo Jun, Qiming Yuan +53
1 version
Preprints & journals
7 papers in the corpus · 2019–2025Beyond Release: Access Considerations for Generative AI Systems2502.16701v2 · Irene Solaiman, Rishi Bommasani, Dan Hendrycks et al.2025 · 0 citationsarXiv
The WMDP Benchmark: Measuring and Reducing Malicious Use With Unlearning2403.03218v7 · Nathaniel Li, Alexander Pan, Anjali Gopal et al.2024 · 13 citationsarXiv
Evaluating Large Language Models Trained on Code2107.03374v2 · Mark Chen, Jerry Tworek, Heewoo Jun et al.2021 · 1,468 citationsarXivon Valency
Extracting Training Data from Large Language Models2012.07805v2 · Nicholas Carlini, Florian Tramer, Eric Wallace et al.2020 · 273 citationsarXiv
Language Models are Few-Shot Learners2005.14165v4 · Tom B. Brown, Benjamin Mann, Nick Ryder et al.2020 · 2,966 citationsarXiv
Toward Trustworthy AI Development: Mechanisms for Supporting Verifiable Claims2004.07213v2 · Miles Brundage, Shahar Avin, Jasmine Wang et al.2020 · 301 citationsarXiv
Release Strategies and the Social Impacts of Language Models1908.09203v2 · Irene Solaiman, Miles Brundage, Jack Clark et al.2019 · 282 citationsarXiv
Career total: 8 works. 7 are in this corpus.
Profile built from the corpus for this byline.
Author records are still filling in while the Hub is in alpha. If this is your page, you'll be able to claim it soon. Spot a mistake? Tell us.