SG

Scott Gray

cs.LGstat.MLcs.CLcs.CVcs.AIcs.CYcs.NAcs.NEcs.SDeess.AS

On Valency

published · living versions
W_3seyyh2p·v1 · currentpublished
Evaluating Large Language Models Trained on Code
with Mark Chen, Jerry Tworek, Heewoo Jun, Qiming Yuan +53
1 version

Preprints & journals

10 papers in the corpus · 2015–2024
Flexpoint: An Adaptive Numerical Format for Efficient Training of Deep Neural Networks1711.02213v2 · Urs Koster, Tristan J. Webb, Xin Wang et al.2017 · 244 citationsarXiv
GPT-4o System Card2410.21276v1 · OpenAI: Aaron Hurst, Adam Lerer, Adam P. Goucher et al.2024 · 165 citationsarXiv
Evaluating Large Language Models Trained on Code2107.03374v2 · Mark Chen, Jerry Tworek, Heewoo Jun et al.2021 · 1,468 citationsarXivon Valency
Dota 2 with Large Scale Deep Reinforcement Learning1912.06680v1 · OpenAI: Christopher Berner, Greg Brockman, Brooke Chan et al.2019 · 1,036 citationsarXiv
Zero-Shot Text-to-Image Generation2102.12092v2 · Aditya Ramesh, Mikhail Pavlov, Gabriel Goh et al.2021 · 1,154 citationsarXiv
Scaling Laws for Autoregressive Generative Modeling2010.14701v2 · Tom Henighan, Jared Kaplan, Mor Katz et al.2020 · 150 citationsarXiv
Language Models are Few-Shot Learners2005.14165v4 · Tom B. Brown, Benjamin Mann, Nick Ryder et al.2020 · 2,966 citationsarXiv
Scaling Laws for Neural Language Models2001.08361v1 · Jared Kaplan, Sam McCandlish, Tom Henighan et al.2020 · 1,513 citationsarXivon Valency
Generating Long Sequences with Sparse Transformers1904.10509v1 · Rewon Child, Scott Gray, Alec Radford et al.2019 · 463 citationsarXiv
Fast Algorithms for Convolutional Neural Networks1509.09308v2 · Andrew Lavin, Scott Gray2015 · 1,013 citationsarXiv
Career total: 13 works. 10 are in this corpus.

Profile built from the corpus for this byline.

Author records are still filling in while the Hub is in alpha. If this is your page, you'll be able to claim it soon. Spot a mistake? Tell us.