NM
Nat McAleese
cs.CLcs.LGcs.AIcs.CRcs.MAcs.SE
On Valency
published · living versionsW_qszyds88·v1 · currentpublished
Competitive Programming with Large Reasoning Models
with OpenAI, :, Ahmed El-Kishky, Alexander Wei +21
1 version
Preprints & journals
10 papers in the corpus · 2021–2025Competitive Programming with Large Reasoning Models2502.06807v2 · OpenAI: Ahmed El-Kishky, Alexander Wei, Andre Saraiva et al.2025 · 6 citationsarXivon Valency
Prover-Verifier Games improve legibility of LLM outputs2407.13692v2 · Jan Hendrik Kirchner, Yining Chen, Harri Edwards et al.2024 · 2 citationsarXiv
LLM Critics Help Catch LLM Bugs2407.00215v1 · Nat McAleese, Rai Michael Pokorny, Juan Felipe Ceron Uribe et al.2024 · 8 citationsarXiv
Fine-Tuning Language Models via Epistemic Neural Networks2211.01568v2 · Ian Osband, Seyed Mohammad Asghari, Benjamin Van Roy et al.2022 · 3 citationsarXiv
Fine-tuning language models to find agreement among humans with diverse preferences2211.15006v1 · Michiel A. Bakker, Martin J. Chadwick, Hannah R. Sheahan et al.2022 · 123 citationsarXiv
Improving alignment of dialogue agents via targeted human judgements2209.14375v1 · Amelia Glaese, Nat McAleese, Maja Trebacz et al.2022 · 132 citationsarXiv
Teaching language models to support answers with verified quotes2203.11147v1 · Jacob Menick, Maja Trebacz, Vladimir Mikulik et al.2022 · 52 citationsarXiv
Red Teaming Language Models with Language Models2202.03286v1 · Ethan Perez, Saffron Huang, Francis Song et al.2022 · 347 citationsarXiv
Scaling Language Models: Methods, Analysis & Insights from Training Gopher2112.11446v2 · Jack W. Rae, Sebastian Borgeaud, Trevor Cai et al.2021 · 245 citationsarXiv
Open-Ended Learning Leads to Generally Capable Agents2107.12808v2 · Open Ended Learning Team, Adam Stooke, Anuj Mahajan et al.2021 · 55 citationsarXiv
Career total: 12 works. 10 are in this corpus.
Profile built from the corpus for this byline.
Author records are still filling in while the Hub is in alpha. If this is your page, you'll be able to claim it soon. Spot a mistake? Tell us.