IS

Ilya Sutskever

cs.LGstat.MLcs.CLcs.AIcs.NEcs.CVcs.SDeess.ASNeural Networks, Computercs.RO

On Valency

published · living versions
W_9msk58k4·v1 · currentpublished
Intriguing properties of neural networks
with Christian Szegedy, Wojciech Zaremba, Joan Bruna, Dumitru Erhan +2
1 version

Preprints & journals

59 papers in the corpus · 2008–2024
Continuous Deep Q-Learning with Model-based Acceleration1603.00748v1 · Shixiang Gu, Timothy Lillicrap, Ilya Sutskever et al.2016 · 320 citationsarXiv
OpenAI o1 System Card2412.16720v2 · OpenAI: Aaron Jaech, Adam Kalai, Adam Lerer et al.2024 · 44 citationsarXiv
GPT-4o System Card2410.21276v1 · OpenAI: Aaron Hurst, Adam Lerer, Adam P. Goucher et al.2024 · 165 citationsarXiv
Scaling and evaluating sparse autoencoders2406.04093v1 · Leo Gao, Tom Dupr'e la Tour, Henk Tillman et al.2024 · 11 citationsarXivon Valency
Weak-to-Strong Generalization: Eliciting Strong Capabilities With Weak Supervision2312.09390v1 · Collin Burns, Pavel Izmailov, Jan Hendrik Kirchner et al.2023 · 25 citationsarXiv
Consistency Models2303.01469v2 · Yang Song, Prafulla Dhariwal, Mark Chen et al.2023 · 23 citationsarXiv
Let's Verify Step by Step2305.20050v1 · Hunter Lightman, Vineet Kosaraju, Yura Burda et al.2023 · 31 citationsarXiv
Robust Speech Recognition via Large-Scale Weak Supervision2212.04356v1 · Alec Radford, Jong Wook Kim, Tao Xu et al.2022 · 1,188 citationsarXiv
GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models2112.10741v3 · Alex Nichol, Prafulla Dhariwal, Aditya Ramesh et al.2021 · 989 citationsarXiv
Formal Mathematics Statement Curriculum Learning2202.01344v1 · Stanislas Polu, Jesse Michael Han, Kunhao Zheng et al.2022 · 24 citationsarXiv
Unsupervised Neural Machine Translation with Generative Language Models Only2110.05448v1 · Jesse Michael Han, Igor Babuschkin, Harrison Edwards et al.2021 · 10 citationsarXiv
Evaluating Large Language Models Trained on Code2107.03374v2 · Mark Chen, Jerry Tworek, Heewoo Jun et al.2021 · 1,468 citationsarXivon Valency
Dota 2 with Large Scale Deep Reinforcement Learning1912.06680v1 · OpenAI: Christopher Berner, Greg Brockman, Brooke Chan et al.2019 · 1,036 citationsarXiv
Zero-Shot Text-to-Image Generation2102.12092v2 · Aditya Ramesh, Mikhail Pavlov, Gabriel Goh et al.2021 · 1,154 citationsarXiv
Learning Transferable Visual Models From Natural Language Supervision2103.00020v1 · Alec Radford, Jong Wook Kim, Chris Hallacy et al.2021 · 5,285 citationsarXiv
Generative Language Modeling for Automated Theorem Proving2009.03393v1 · Stanislas Polu, Ilya Sutskever2020 · 42 citationsarXiv
Language Models are Few-Shot Learners2005.14165v4 · Tom B. Brown, Benjamin Mann, Nick Ryder et al.2020 · 2,966 citationsarXiv
Jukebox: A Generative Model for Music2005.00341v1 · Prafulla Dhariwal, Heewoo Jun, Christine Payne et al.2020 · 102 citationsarXiv
Deep Double Descent: Where Bigger Models and More Data Hurt1912.02292v1 · Preetum Nakkiran, Gal Kaplun, Yamini Bansal et al.2019 · 503 citationsarXiv
Third-Person Imitation Learning1703.01703v2 · Bradly C. Stadie, Pieter Abbeel, Ilya Sutskever2017 · 146 citationsarXiv
Generating Long Sequences with Sparse Transformers1904.10509v1 · Rewon Child, Scott Gray, Alec Radford et al.2019 · 463 citationsarXiv
Some Considerations on Learning to Explore via Meta-Reinforcement Learning1803.01118v2 · Bradly C. Stadie, Ge Yang, Rein Houthooft et al.2018 · 66 citationsarXiv
GamePad: A Learning Environment for Theorem Proving1806.00608v2 · Daniel Huang, Prafulla Dhariwal, Dawn Song et al.2018 · 58 citationsarXiv
FFJORD: Free-form Continuous Dynamics for Scalable Reversible Generative Models1810.01367v3 · Will Grathwohl, Ricky T. Q. Chen, Jesse Bettencourt et al.2018 · 439 citationsarXiv
Emergent Complexity via Multi-Agent Competition1710.03748v3 · Trapit Bansal, Jakub Pachocki, Szymon Sidor et al.2017 · 144 citationsarXiv
Continuous Adaptation via Meta-Learning in Nonstationary and Competitive Environments1710.03641v2 · Maruan Al-Shedivat, Trapit Bansal, Yuri Burda et al.2017 · 246 citationsarXiv
One-Shot Imitation Learning1703.07326v3 · Yan Duan, Marcin Andrychowicz, Bradly C. Stadie et al.2017 · 226 citationsarXiv
An online sequence-to-sequence model for noisy speech recognition1706.06428v1 · Chung-Cheng Chiu, Dieterich Lawson, Yuping Luo et al.2017 · 3 citationsarXiv
Learning to Generate Reviews and Discovering Sentiment1704.01444v2 · Alec Radford, Rafal Jozefowicz, Ilya Sutskever2017 · 346 citationsarXiv
Variational Lossy Autoencoder1611.02731v2 · Xi Chen, Diederik P. Kingma, Tim Salimans et al.2016 · 241 citationsarXiv
Improving Variational Inference with Inverse Autoregressive Flow1606.04934v2 · Diederik P. Kingma, Tim Salimans, Rafal Jozefowicz et al.2016 · 180 citationsarXiv
RL$^2$: Fast Reinforcement Learning via Slow Reinforcement Learning1611.02779v2 · Yan Duan, John Schulman, Xi Chen et al.2016 · 479 citationsarXiv
Extensions and Limitations of the Neural GPU1611.00736v2 · Eric Price, Wojciech Zaremba, Ilya Sutskever2016 · 11 citationsarXiv
A Neural Transducer1511.04868v4 · Navdeep Jaitly, David Sussillo, Quoc V. Le et al.2015 · 36 citationsarXiv
Neural Programmer: Inducing Latent Programs with Gradient Descent1511.04834v3 · Arvind Neelakantan, Quoc V. Le, Ilya Sutskever2015 · 71 citationsarXiv
Learning Online Alignments with Continuous Rewards Policy Gradient1608.01281v1 · Yuping Luo, Chung-Cheng Chiu, Navdeep Jaitly et al.2016 · 40 citationsarXiv
TensorFlow: Large-Scale Machine Learning on Heterogeneous Distributed Systems1603.04467v2 · Mart'in Abadi, Ashish Agarwal, Paul Barham et al.2016 · 9,713 citationsarXivon Valency
Neural GPUs Learn Algorithms1511.08228v3 · \Lukasz Kaiser, Ilya Sutskever2015 · 225 citationsarXiv
Multi-task Sequence to Sequence Learning1511.06114v4 · Minh-Thang Luong, Quoc V. Le, Ilya Sutskever et al.2015 · 65 citationsarXiv
MuProp: Unbiased Backpropagation for Stochastic Neural Networks1511.05176v3 · Shixiang Gu, Sergey Levine, Ilya Sutskever et al.2015 · 30 citationsarXiv
Neural Random-Access Machines1511.06392v3 · Karol Kurach, Marcin Andrychowicz, Ilya Sutskever2015 · 14 citationsarXiv
Mastering the game of Go with deep neural networks and tree search.26819042 · Silver, David, Huang, Aja, Maddison, Chris J et al.2016 · 16,149 citationsNature. 2016;529(7587):484-9
Reinforcement Learning Neural Turing Machines - Revised1505.00521v3 · Wojciech Zaremba, Ilya Sutskever2015 · 112 citationsarXiv
Towards Principled Unsupervised Learning1511.06440v2 · Ilya Sutskever, Rafal Jozefowicz, Karol Gregor et al.2015 · 20 citationsarXiv
Learning to Execute1410.4615v3 · Wojciech Zaremba, Ilya Sutskever2014 · 305 citationsarXiv
Adding Gradient Noise Improves Learning for Very Deep Networks1511.06807v1 · Arvind Neelakantan, Luke Vilnis, Quoc V. Le et al.2015 · 258 citationsarXiv
Grammar as a Foreign Language1412.7449v3 · Oriol Vinyals, Lukasz Kaiser, Terry Koo et al.2014 · 399 citationsarXiv
Addressing the Rare Word Problem in Neural Machine Translation1410.8206v4 · Minh-Thang Luong, Ilya Sutskever, Quoc V. Le et al.2014 · 736 citationsarXiv
Recurrent Neural Network Regularization1409.2329v5 · Wojciech Zaremba, Ilya Sutskever, Oriol Vinyals2014 · 2,267 citationsarXivon Valency
Career total: 98 works. 59 are in this corpus.Showing the 50 most recent.

Profile built from the corpus for this byline.

Author records are still filling in while the Hub is in alpha. If this is your page, you'll be able to claim it soon. Spot a mistake? Tell us.