AR

Alec Radford

cs.LGcs.CLcs.AIcs.CVstat.MLcs.CYcs.SDeess.AScs.NE

On Valency

published · living versions
W_b2ety6g8·v1 · currentpublished
Scaling and evaluating sparse autoencoders
with Leo Gao, Tom Dupré la Tour, Henk Tillman, Gabriel Goh +4
1 version

Preprints & journals

23 papers in the corpus · 2015–2026
Learning a Generative Meta-Model of LLM Activations2602.06964v1 · Grace Luo, Jiahai Feng, Trevor Darrell et al.2026 · 0 citationsarXiv
Shaping capabilities with token-level data filtering2601.21571v2 · Neil Rathi, Alec Radford2026 · 0 citationsarXiv
GPT-4o System Card2410.21276v1 · OpenAI: Aaron Hurst, Adam Lerer, Adam P. Goucher et al.2024 · 165 citationsarXiv
Scaling and evaluating sparse autoencoders2406.04093v1 · Leo Gao, Tom Dupr'e la Tour, Henk Tillman et al.2024 · 11 citationsarXivon Valency
Robust Speech Recognition via Large-Scale Weak Supervision2212.04356v1 · Alec Radford, Jong Wook Kim, Tao Xu et al.2022 · 1,188 citationsarXiv
Learning to summarize from human feedback2009.01325v3 · Nisan Stiennon, Long Ouyang, Jeff Wu et al.2020 · 36 citationsarXiv
Text and Code Embeddings by Contrastive Pre-Training2201.10005v1 · Arvind Neelakantan, Tao Xu, Raul Puri et al.2022 · 149 citationsarXiv
Unsupervised Neural Machine Translation with Generative Language Models Only2110.05448v1 · Jesse Michael Han, Igor Babuschkin, Harrison Edwards et al.2021 · 10 citationsarXiv
Evaluating CLIP: Towards Characterization of Broader Capabilities and Downstream Implications2108.02818v1 · Sandhini Agarwal, Gretchen Krueger, Jack Clark et al.2021 · 36 citationsarXiv
Evaluating Large Language Models Trained on Code2107.03374v2 · Mark Chen, Jerry Tworek, Heewoo Jun et al.2021 · 1,468 citationsarXivon Valency
Zero-Shot Text-to-Image Generation2102.12092v2 · Aditya Ramesh, Mikhail Pavlov, Gabriel Goh et al.2021 · 1,154 citationsarXiv
Learning Transferable Visual Models From Natural Language Supervision2103.00020v1 · Alec Radford, Jong Wook Kim, Chris Hallacy et al.2021 · 5,285 citationsarXiv
Scaling Laws for Autoregressive Generative Modeling2010.14701v2 · Tom Henighan, Jared Kaplan, Mor Katz et al.2020 · 150 citationsarXiv
Language Models are Few-Shot Learners2005.14165v4 · Tom B. Brown, Benjamin Mann, Nick Ryder et al.2020 · 2,966 citationsarXiv
Jukebox: A Generative Model for Music2005.00341v1 · Prafulla Dhariwal, Heewoo Jun, Christine Payne et al.2020 · 102 citationsarXiv
Scaling Laws for Neural Language Models2001.08361v1 · Jared Kaplan, Sam McCandlish, Tom Henighan et al.2020 · 1,513 citationsarXivon Valency
Fine-Tuning Language Models from Human Preferences1909.08593v2 · Daniel M. Ziegler, Nisan Stiennon, Jeffrey Wu et al.2019 · 385 citationsarXiv
Release Strategies and the Social Impacts of Language Models1908.09203v2 · Irene Solaiman, Miles Brundage, Jack Clark et al.2019 · 282 citationsarXiv
Generating Long Sequences with Sparse Transformers1904.10509v1 · Rewon Child, Scott Gray, Alec Radford et al.2019 · 463 citationsarXiv
Improving GANs Using Optimal Transport1803.05573v1 · Tim Salimans, Han Zhang, Alec Radford et al.2018 · 45 citationsarXiv
Learning to Generate Reviews and Discovering Sentiment1704.01444v2 · Alec Radford, Rafal Jozefowicz, Ilya Sutskever2017 · 346 citationsarXiv
Improved Techniques for Training GANs1606.03498v1 · Tim Salimans, Ian Goodfellow, Wojciech Zaremba et al.2016 · 1,344 citationsarXiv
Unsupervised Representation Learning with Deep Convolutional Generative Adversarial Networks1511.06434v2 · Alec Radford, Luke Metz, Soumith Chintala2015 · 12,526 citationsarXiv
Career total: 28 works. 23 are in this corpus.

Profile built from the corpus for this byline.

Author records are still filling in while the Hub is in alpha. If this is your page, you'll be able to claim it soon. Spot a mistake? Tell us.