RC
Rewon Child
cs.CLcs.LGstat.MLcs.AIcs.CVcs.NE
On Valency
published · living versionsW_tkk5bk3v·v1 · currentpublished
Scaling Laws for Neural Language Models
with Jared Kaplan, Sam McCandlish, Tom Henighan, Tom B. Brown +5
1 version
Preprints & journals
10 papers in the corpus · 2016–2022PaLM: Scaling Language Modeling with Pathways2204.02311v5 · Aakanksha Chowdhery, Sharan Narang, Jacob Devlin et al.2022 · 2,138 citationsarXiv
Using DeepSpeed and Megatron to Train Megatron-Turing NLG 530B, A Large-Scale Generative Language Model2201.11990v3 · Shaden Smith, Mostofa Patwary, Brandon Norick et al.2022 · 296 citationsarXiv
Very Deep VAEs Generalize Autoregressive Models and Can Outperform Them on Images2011.10650v2 · Rewon Child2020 · 118 citationsarXiv
Language Models are Few-Shot Learners2005.14165v4 · Tom B. Brown, Benjamin Mann, Nick Ryder et al.2020 · 2,966 citationsarXiv
Scaling Laws for Neural Language Models2001.08361v1 · Jared Kaplan, Sam McCandlish, Tom Henighan et al.2020 · 1,513 citationsarXivon Valency
Generating Long Sequences with Sparse Transformers1904.10509v1 · Rewon Child, Scott Gray, Alec Radford et al.2019 · 463 citationsarXiv
Exploring Neural Transducers for End-to-End Speech Recognition1707.07413v1 · Eric Battenberg, Jitong Chen, Rewon Child et al.2017 · 257 citationsarXiv
Convolutional Recurrent Neural Networks for Small-Footprint Keyword Spotting1703.05390v3 · Sercan O. Arik, Markus Kliegl, Rewon Child et al.2017 · 156 citationsarXiv
Reducing Bias in Production Speech Models1705.04400v1 · Eric Battenberg, Rewon Child, Adam Coates et al.2017 · 11 citationsarXiv
Active Learning for Speech Recognition: the Power of Gradients1612.03226v1 · Jiaji Huang, Rewon Child, Vinay Rao et al.2016 · 48 citationsarXiv
Career total: 15 works. 10 are in this corpus.
Profile built from the corpus for this byline.
Author records are still filling in while the Hub is in alpha. If this is your page, you'll be able to claim it soon. Spot a mistake? Tell us.