WF

William Fedus

cs.LGcs.CLstat.MLcs.AIcs.CVcancercomputer visioncs.CYcs.ITcs.SI

On Valency

published · living versions
W_5vz5yxu5·v1 · currentpublished
Many Paths to Equilibrium: GANs Do Not Need to Decrease a Divergence At Every Step
with Mihaela Rosca, Balaji Lakshminarayanan, Andrew M. Dai, Shakir Mohamed +1
1 version

Preprints & journals

25 papers in the corpus · 2017–2025
BrowseComp: A Simple Yet Challenging Benchmark for Browsing Agents2504.12516v1 · Jason Wei, Zhiqing Sun, Spencer Papay et al.2025 · 2 citationsarXiv
Measuring short-form factuality in large language models2411.04368v1 · Jason Wei, Nguyen Karina, Hyung Won Chung et al.2024 · 13 citationsarXiv
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models2206.04615v3 · Aarohi Srivastava, Abhinav Rastogi, Abhishek Rao et al.2022 · 565 citationsTransactions on Machine Learning Research, May/2022, https://openreview.net/forum?id=uyTL5Bvosj
Scaling Instruction-Finetuned Language Models2210.11416v5 · Hyung Won Chung, Le Hou, Shayne Longpre et al.2022 · 1,174 citationsarXiv
Emergent Abilities of Large Language Models2206.07682v2 · Jason Wei, Yi Tay, Rishi Bommasani et al.2022 · 1,047 citationsarXiv
A Review of Sparse Expert Models in Deep Learning2209.01667v1 · William Fedus, Jeff Dean, Barret Zoph2022 · 33 citationsarXiv
Scaling Laws vs Model Architectures: How does Inductive Bias Influence Scaling?2207.10551v1 · Yi Tay, Mostafa Dehghani, Samira Abnar et al.2022 · 42 citationsarXiv
Switch Transformers: Scaling to Trillion Parameter Models with Simple and Efficient Sparsity2101.03961v3 · William Fedus, Barret Zoph, Noam Shazeer2021 · 918 citationsarXiv
ST-MoE: Designing Stable and Transferable Sparse Expert Models2202.08906v2 · Barret Zoph, Irwan Bello, Sameer Kumar et al.2022 · 52 citationsarXiv
Scale Efficiently: Insights from Pre-training and Fine-tuning Transformers2109.10686v2 · Yi Tay, Mostafa Dehghani, Jinfeng Rao et al.2021 · 56 citationsarXiv
Benchmarking Bonus-Based Exploration Methods on the Arcade Learning Environment1908.02388v3 · Adrien Ali Taiga, William Fedus, Marlos C. Machado et al.2019 · 23 citationsarXiv
On Bonus-Based Exploration Methods in the Arcade Learning Environment2109.11052v1 · Adrien Ali Taiga, William Fedus, Marlos C. Machado et al.2021 · 26 citationsPublished as a conference paper at ICLR 2020
Revisiting ResNets: Improved Training and Scaling Strategies2103.07579v1 · Irwan Bello, William Fedus, Xianzhi Du et al.2021 · 219 citationsarXiv
Revisiting Fundamentals of Experience Replay2007.06700v1 · William Fedus, Prajit Ramachandran, Rishabh Agarwal et al.2020 · 79 citationsarXiv
On Catastrophic Interference in Atari 2600 Games2002.12499v2 · William Fedus, Dibya Ghosh, John D. Martin et al.2020 · 3 citationsarXiv
Language GANs Falling Short1811.02549v6 · Massimo Caccia, Lucas Caccia, William Fedus et al.2018 · 79 citationsICLR 2020 - Proceedings of the Seventh International Conference on Learning Representation
Algorithmic Improvements for Deep Reinforcement Learning applied to Interactive Fiction1911.12511v1 · Vishal Jain, William Fedus, Hugo Larochelle et al.2019 · 22 citationsarXiv
Hyperbolic Discounting and Learning over Multiple Horizons1902.06865v3 · William Fedus, Carles Gelada, Yoshua Bengio et al.2019 · 49 citationsarXiv
In Silico Labeling: Predicting Fluorescent Labels in Unlabeled Images.29656897 · Christiansen, Eric M, Yang, Samuel J, Ando, D Michael et al.2019 · 693 citationsCell. 2018;173(3):792-803.e19
Recall Traces: Backtracking Models for Efficient Reinforcement Learning1804.00379v2 · Anirudh Goyal, Philemon Brakel, William Fedus et al.2018 · 23 citationsarXiv
Deep Graph Infomax1809.10341v2 · Petar Velickovi'c, William Fedus, William L. Hamilton et al.2018 · 82 citationsarXiv
MaskGAN: Better Text Generation via Filling in the______1801.07736v3 · William Fedus, Ian Goodfellow, Andrew M. Dai2018 · 354 citationsarXiv
Disentangling the independently controllable factors of variation by interacting with the world1802.09484v1 · Valentin Thomas, Emmanuel Bengio, William Fedus et al.2018 · 44 citationsarXiv
Many Paths to Equilibrium: GANs Do Not Need to Decrease a Divergence At Every Step1710.08446v3 · William Fedus, Mihaela Rosca, Balaji Lakshminarayanan et al.2017 · 159 citationsarXivon Valency
Career total: 32 works. 25 are in this corpus.

Profile built from the corpus for this byline.

Author records are still filling in while the Hub is in alpha. If this is your page, you'll be able to claim it soon. Spot a mistake? Tell us.