DA

Dario Amodei

cs.LGcs.CLcs.AIstat.MLq-bio.NCcond-mat.stat-mechcs.CYcond-mat.dis-nncs.HCphysics.bio-ph

On Valency

published · living versions
W_6uv2jkb5·v1 · currentpublished
Thermodynamics for a network of neurons: Signatures of criticality
with Gasper Tkacik, Thierry Mora, Olivier Marre, Michael J. Berry +1
1 version

Preprints & journals

43 papers in the corpus · 2003–2023
Thermodynamics for a network of neurons: Signatures of criticality1407.5946v1 · Gasper Tkacik, Thierry Mora, Olivier Marre et al.2014 · 23 citationsProceedings of the National Academy of Sciences (USA)112, 11508-11513 (2015)on Valency
The Malicious Use of Artificial Intelligence: Forecasting, Prevention, and Mitigation1802.07228v2 · Miles Brundage, Shahar Avin, Jack Clark et al.2018 · 495 citationsarXivon Valency
The Capacity for Moral Self-Correction in Large Language Models2302.07459v2 · Deep Ganguli, Amanda Askell, Nicholas Schiefer et al.2023 · 53 citationsarXiv
Deep reinforcement learning from human preferences1706.03741v4 · Paul Christiano, Jan Leike, Tom B. Brown et al.2017 · 500 citationsarXiv
Discovering Language Model Behaviors with Model-Written Evaluations2212.09251v1 · Ethan Perez, Sam Ringer, Kamil\.e Lukosi\=ut\.e et al.2022 · 224 citationsarXivon Valency
Constitutional AI: Harmlessness from AI Feedback2212.08073v1 · Yuntao Bai, Saurav Kadavath, Sandipan Kundu et al.2022 · 322 citationsarXivon Valency
Red Teaming Language Models to Reduce Harms: Methods, Scaling Behaviors, and Lessons Learned2209.07858v2 · Deep Ganguli, Liane Lovitt, Jackson Kernion et al.2022 · 119 citationsarXivon Valency
Language Models (Mostly) Know What They Know2207.05221v4 · Saurav Kadavath, Tom Conerly, Amanda Askell et al.2022 · 169 citationsarXivon Valency
Measuring Progress on Scalable Oversight for Large Language Models2211.03540v2 · Samuel R. Bowman, Jeeyoon Hyun, Ethan Perez et al.2022 · 34 citationsarXiv
Predictability and Surprise in Large Generative Models2202.07785v2 · Deep Ganguli, Danny Hernandez, Liane Lovitt et al.2022 · 204 citationsarXiv
In-context Learning and Induction Heads2209.11895v1 · Catherine Olsson, Nelson Elhage, Neel Nanda et al.2022 · 87 citationsarXivon Valency
Toy Models of Superposition2209.10652v1 · Nelson Elhage, Tristan Hume, Catherine Olsson et al.2022 · 49 citationsarXiv
Scaling Laws and Interpretability of Learning from Repeated Data2205.10487v1 · Danny Hernandez, Tom Brown, Tom Conerly et al.2022 · 22 citationsarXiv
Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback2204.05862v1 · Yuntao Bai, Andy Jones, Kamal Ndousse et al.2022 · 389 citationsarXiv
Learning to summarize from human feedback2009.01325v3 · Nisan Stiennon, Long Ouyang, Jeff Wu et al.2020 · 36 citationsarXiv
A General Language Assistant as a Laboratory for Alignment2112.00861v3 · Amanda Askell, Yuntao Bai, Anna Chen et al.2021 · 27 citationsarXiv
Evaluating Large Language Models Trained on Code2107.03374v2 · Mark Chen, Jerry Tworek, Heewoo Jun et al.2021 · 1,468 citationsarXivon Valency
Scaling Laws for Autoregressive Generative Modeling2010.14701v2 · Tom Henighan, Jared Kaplan, Mor Katz et al.2020 · 150 citationsarXiv
Language Models are Few-Shot Learners2005.14165v4 · Tom B. Brown, Benjamin Mann, Nick Ryder et al.2020 · 2,966 citationsarXiv
Physical Principles for Scalable Neural Recording1306.5709v7 · Adam H. Marblestone, Bradley M. Zamft, Yael G. Maguire et al.2013 · 267 citationsarXiv
Scaling Laws for Neural Language Models2001.08361v1 · Jared Kaplan, Sam McCandlish, Tom Henighan et al.2020 · 1,513 citationsarXivon Valency
Fine-Tuning Language Models from Human Preferences1909.08593v2 · Daniel M. Ziegler, Nisan Stiennon, Jeffrey Wu et al.2019 · 385 citationsarXiv
Improving Precursor Selectivity in Data-Independent Acquisition Using Overlapping Windows.30671891 · Amodei, Dario, Egertson, Jarrett, MacLean, Brendan X et al.2019 · 170 citationsJournal of the American Society for Mass Spectrometry. 2019;30(4):669-684
An Empirical Model of Large-Batch Training1812.06162v1 · Sam McCandlish, Jared Kaplan, Dario Amodei et al.2018 · 129 citationsarXiv
Reward learning from human preferences and demonstrations in Atari1811.06521v1 · Borja Ibarz, Jan Leike, Tobias Pohlen et al.2018 · 39 citationsarXiv
AI safety via debate1805.00899v2 · Geoffrey Irving, Paul Christiano, Dario Amodei2018 · 29 citationsarXiv
Supervising strong learners by amplifying weak experts1810.08575v1 · Paul Christiano, Buck Shlegeris, Dario Amodei2018 · 26 citationsarXiv
Variational Option Discovery Algorithms1807.10299v1 · Joshua Achiam, Harrison Edwards, Dario Amodei et al.2018 · 78 citationsarXiv
Learning a Natural Language Interface with Neural Programmer1611.08945v4 · Arvind Neelakantan, Quoc V. Le, Martin Abadi et al.2016 · 33 citationsarXiv
Identification of a Set of Conserved Eukaryotic Internal Retention Time Standards for Data-independent Acquisition Mass Spectrometry.26199342 · Parker, Sarah J, Rost, Hannes, Rosenberger, George et al.2016 · 100 citationsMolecular & cellular proteomics : MCP. 2015;14(10):2800-13
Thermodynamics and signatures of criticality in a network of neurons.26330611 · Tkačik, Gašper, Mora, Thierry, Marre, Olivier et al.2015 · 275 citationsProceedings of the National Academy of Sciences of the United States of America. 2015;112(37):11508-13
Deep Speech 2: End-to-End Speech Recognition in English and Mandarin1512.02595v1 · Dario Amodei, Rishita Anubhai, Eric Battenberg et al.2015 · 2,265 citationsarXiv
Building high-quality assay libraries for targeted analysis of SWATH MS data.25675208 · Schubert, Olga T, Gillet, Ludovic C, Collins, Ben C et al.2015 · 339 citationsNature protocols. 2015;10(3):426-41
The simplest maximum entropy model for collective behavior in a neural network1207.6319v1 · Gasper Tkacik, Olivier Marre, Thierry Mora et al.2012 · 148 citationsarXiv
Searching for collective behavior in a network of real neurons1306.3061v1 · Gasper Tkacik, Olivier Marre, Dario Amodei et al.2013 · 278 citationsPLOS Comput Biol 10 (2014): e1003408on Valency
Characterizing deformability and surface friction of cancer cells.23610435 · Byun, Sangwon, Son, Sungmin, Amodei, Dario et al.2013 · 373 citationsProceedings of the National Academy of Sciences of the United States of America. 2013;110(19):7580-5
Low error discrimination using a correlated population code.22539825 · Schwartz, Greg, Macke, Jakob, Amodei, Dario et al.2013 · 27 citationsJournal of neurophysiology. 2012;108(4):1069-88
A cross-platform toolkit for mass spectrometry and proteomics.23051804 · Chambers, Matthew C, Maclean, Brendan, Burke, Robert et al.2013 · 4,470 citationsNature biotechnology. 2012;30(10):918-20
Mapping a complete neural population in the retina.23100409 · Marre, Olivier, Amodei, Dario, Deshmukh, Nikhil et al.2013 · 193 citationsThe Journal of neuroscience : the official journal of the Society for Neuroscience. 2012;32(43):14859-73
Computation of uniform wave forms using complex rays.16605695 · Amodei, Dario, Keers, Henk, Vasco, Don et al.2006 · 21 citationsPhysical review. E, Statistical, nonlinear, and soft matter physics. 2006;73(3 Pt 2):036704
Mirrors with regular hexagonal segments.12962392 · Amodei, Dario, Padin, Stephen2003 · 5 citationsApplied optics. 2003;42(25):5130-5
Career total: 48 works. 43 are in this corpus.Showing the 41 most recent.

Profile built from the corpus for this byline.

Author records are still filling in while the Hub is in alpha. If this is your page, you'll be able to claim it soon. Spot a mistake? Tell us.