CR
Christopher Ré
cs.LGstat.MLcs.AIcs.CLcs.DBcs.CVmath.OCcs.DCcs.DScs.IR
On Valency
published · living versionsW_u7fa5rax·v1 · currentpublished
MLSys: The New Frontier of Machine Learning Systems
with Alexander Ratner, Dan Alistarh, Gustavo Alonso, David G. Andersen +64
1 version
Preprints & journals
187 papers in the corpus · 2011–2025Intelligence per Watt: Measuring Intelligence Efficiency of Local AI2511.07885v7 · Jon Saad-Falcon, Avanika Narayan, Hakki Orhun Akengin et al.2025 · 0 citationsarXiv
An Approximate, Efficient Solver for LP Rounding1311.2661v2 · Srikrishna Sridhar, Victor Bittorf, Ji Liu et al.2013 · 1 citationarXiv
Accelerated Stochastic Power Iteration1707.02670v1 · Christopher De Sa, Bryan He, Ioannis Mitliagkas et al.2017 · 20 citationsarXiv
Cartridges: Lightweight and general-purpose long context representations via self-study2506.06266v3 · Sabri Eyuboglu, Ryan Ehrlich, Simran Arora et al.2025 · 0 citationsarXiv
Archon: An Architecture Search Framework for Inference-Time Techniques2409.15254v6 · Jon Saad-Falcon, Adrian Gamarra Lafuente, Shlok Natarajan et al.2024 · 0 citationsarXiv
Precise High-Dimensional Asymptotics for Quantifying Heterogeneous Transfers2010.11750v5 · Fan Yang, Hongyang R. Zhang, Sen Wu et al.2020 · 1 citationJ. Mach. Learn. Res. 26 (2025)
HMAR: Efficient Hierarchical Masked Auto-Regressive Image Generation2506.04421v1 · Hermann Kumbong, Xian Liu, Tsung-Yi Lin et al.2025 · 1 citationarXiv
Context Clues: Evaluating Long Context Models for Clinical Prediction Tasks on EHRs2412.16178v2 · Michael Wornow, Suhana Bedi, Miguel Angel Fuentes Hernandez et al.2024 · 3 citationsarXiv
Restructuring Vector Quantization with the Rotation Trick2410.06424v2 · Christopher Fifty, Ronald G. Junkins, Dennis Duan et al.2024 · 1 citationarXiv
Language Models Enable Simple Systems for Generating Structured Views of Heterogeneous Data Lakes2304.09433v3 · Simran Arora, Brandon Yang, Sabri Eyuboglu et al.2023 · 75 citationsarXiv
Simple linear attention language models balance the recall-throughput tradeoff2402.18668v2 · Simran Arora, Sabri Eyuboglu, Michael Zhang et al.2024 · 4 citationsarXiv
LoLCATs: On Low-Rank Linearizing of Large Language Models2410.10254v3 · Michael Zhang, Simran Arora, Rahul Chalamala et al.2024 · 1 citationarXiv
Systems and Algorithms for Convolutional Multi-Hybrid Language Models at Scale2503.01868v1 · Jerome Ku, Eric Nguyen, David W. Romero et al.2025 · 5 citationsarXiv
Minions: Cost-efficient Collaboration Between On-device and Cloud Language Models2502.15964v1 · Avanika Narayan, Dan Biderman, Sabri Eyuboglu et al.2025 · 1 citationarXiv
KernelBench: Can LLMs Write Efficient GPU Kernels?2502.10517v1 · Anne Ouyang, Simon Guo, Simran Arora et al.2025 · 1 citationarXiv
CodeMonkeys: Scaling Test-Time Compute for Software Engineering2501.14723v2 · Ryan Ehrlich, Bradley Brown, Jordan Juravsky et al.2025 · 0 citationsarXiv
Large Language Monkeys: Scaling Inference Compute with Repeated Sampling2407.21787v3 · Bradley Brown, Jordan Juravsky, Ryan Ehrlich et al.2024 · 12 citationsarXiv
Correct-N-Contrast: A Contrastive Approach for Improving Robustness to Spurious Correlations2203.01517v2 · Michael Zhang, Nimit S. Sohoni, Hongyang R. Zhang et al.2022 · 14 citationsarXiv
Smoothie: Label Free Language Model Routing2412.04692v1 · Neel Guha, Mayee F. Chen, Trevor Chow et al.2024 · 0 citationsarXiv
Scaling Laws for Precision2411.04330v2 · Tanishq Kumar, Zachary Ankner, Benjamin F. Spector et al.2024 · 0 citationsarXiv
Benchmarking and Building Long-Context Retrieval Models with LoCo and M2-BERT2402.07440v3 · Jon Saad-Falcon, Daniel Y. Fu, Simran Arora et al.2024 · 1 citationarXiv
ThunderKittens: Simple, Fast, and Adorable AI Kernels2410.20399v1 · Benjamin F. Spector, Simran Arora, Aaryan Singhal et al.2024 · 0 citationsarXiv
Automated Rewards via LLM-Generated Progress Functions2410.09187v2 · Vishnu Sarukkai, Brennan Shacklett, Zander Majercik et al.2024 · 0 citationsarXiv
Cookbook: A framework for improving LLM generative abilities via programmatic data generating templates2410.05224v1 · Avanika Narayan, Mayee F. Chen, Kush Bhatia et al.2024 · 0 citationsarXiv
Medical device surveillance with electronic health records.31583282 · Callahan, Alison, Fries, Jason A, Ré, Christopher et al.2024 · 81 citationsNPJ digital medicine. 2019;2:94
Mechanistic Design and Scaling of Hybrid Architectures2403.17844v2 · Michael Poli, Armin W Thomas, Eric Nguyen et al.2024 · 4 citationsarXiv
Slice-based Learning: A Programming Model for Residual Learning in Critical Data Slices.31871391 · Chen, Vincent S, Wu, Sen, Weng, Zhenzhen et al.2024 · 42 citationsAdvances in neural information processing systems. 2019;32:9392-9402
Training Classifiers with Natural Language Explanations.31130772 · Hancock, Braden, Bringmann, Martin, Varma, Paroma et al.2024 · 120 citationsProceedings of the conference. Association for Computational Linguistics. Meeting. 2018;2018:1884-1895
Representation Tradeoffs for Hyperbolic Embeddings.31131375 · De Sa, Christopher, Gu, Albert, Ré, Christopher et al.2024 · 273 citationsProceedings of machine learning research. 2018;80:4460-4469
ShortFuse: Biomedical Time Series Representations in the Presence of Structured Information.30882086 · Fiterau, Madalina, Bhooshan, Suvrat, Fries, Jason et al.2024 · 2 citationsProceedings of machine learning research. 2017;68:59-74
Just read twice: closing the recall gap for recurrent language models2407.05483v1 · Simran Arora, Aman Timalsina, Aaryan Singhal et al.2024 · 1 citationarXiv
State-Free Inference of State-Space Models: The Transfer Function Approach2405.06147v2 · Rom N. Parnichkun, Stefano Massaroli, Alessandro Moro et al.2024 · 1 citationarXiv
Hydragen: High-Throughput LLM Inference with Shared Prefixes2402.05099v2 · Jordan Juravsky, Bradley Brown, Ryan Ehrlich et al.2024 · 1 citationarXiv
Automating the Enterprise with Foundation Models2405.03710v1 · Michael Wornow, Avanika Narayan, Krista Opsahl-Ong et al.2024 · 13 citationsarXiv
Context-Aware Meta-Learning2310.10971v2 · Christopher Fifty, Dennis Duan, Ronald G. Junkins et al.2023 · 3 citationsarXiv
Inferring Generative Model Structure with Static Analysis.29391769 · Varma, Paroma, He, Bryan, Bajaj, Payal et al.2024 · 46 citationsAdvances in neural information processing systems. 2017;30:239-249
Gaussian Quadrature for Kernel Features.29398882 · Dao, Tri, De Sa, Christopher, Ré, Christopher2024 · 33 citationsAdvances in neural information processing systems. 2017;30:6109-6119
Extracting Databases from Dark Data with DeepDive.28316365 · Zhang, Ce, Shin, Jaeho, Ré, Christopher et al.2024 · 46 citationsProceedings. ACM-SIGMOD International Conference on Management of Data. 2016;2016:847-859
Weighted SGD for ℓ p Regression with Randomized Preconditioning.29782626 · Yang, Jiyan, Chow, Yin-Lam, Ré, Christopher et al.2024 · 18 citationsProceedings of the ... Annual ACM-SIAM Symposium on Discrete Algorithms. ACM-SIAM Symposium on Discrete Algorithms. 2016;2016:558-569
The Hedgehog & the Porcupine: Expressive Linear Attentions with Softmax Mimicry2402.04347v1 · Michael Zhang, Kush Bhatia, Hermann Kumbong et al.2024 · 3 citationsarXiv
H$_2$O: Heavy-Hitter Oracle for Efficient Generative Inference of Large Language Models2306.14048v3 · Zhenyu Zhang, Ying Sheng, Tianyi Zhou et al.2023 · 28 citationsarXiv
Zoology: Measuring and Improving Recall in Efficient Language Models2312.04927v1 · Simran Arora, Sabri Eyuboglu, Aman Timalsina et al.2023 · 1 citationarXiv
HyenaDNA: Long-Range Genomic Sequence Modeling at Single Nucleotide Resolution2306.15794v2 · Eric Nguyen, Michael Poli, Marjan Faizi et al.2023 · 215 citationsarXiv
FlashFFTConv: Efficient Convolutions for Long Sequences with Tensor Cores2311.05908v1 · Daniel Y. Fu, Hermann Kumbong, Eric Nguyen et al.2023 · 1 citationarXiv
Deja Vu: Contextual Sparsity for Efficient LLMs at Inference Time2310.17157v1 · Zichang Liu, Jue Wang, Tri Dao et al.2023 · 19 citationsProceedings of the 40th International Conference on Machine Learning, 2023, 919
Accelerated Stochastic Power Iteration.31187095 · De Sa, Christopher, He, Bryan, Mitliagkas, Ioannis et al.2023 · 20 citationsProceedings of machine learning research. 2018;84:58-67
Learning Compressed Transforms with Low Displacement Rank.31130799 · Thomas, Anna T, Gu, Albert, Dao, Tri et al.2023 · 3 citationsAdvances in neural information processing systems. 2018;2018:9052-9060
Learning the Structure of Generative Models without Labeled Data.30882087 · Bach, Stephen H, He, Bryan, Ratner, Alexander et al.2023 · 113 citationsProceedings of machine learning research. 2017;70:273-82
Holistic Evaluation of Language Models2211.09110v2 · Percy Liang, Rishi Bommasani, Tony Lee et al.2022 · 584 citationsPublished in Transactions on Machine Learning Research (TMLR), 2023
Collage Diffusion2303.00262v2 · Vishnu Sarukkai, Linden Li, Arden Ma et al.2023 · 0 citationsarXiv
Career total: 350 works. 187 are in this corpus.Showing the 50 most recent.
Profile built from the corpus for this byline.
Author records are still filling in while the Hub is in alpha. If this is your page, you'll be able to claim it soon. Spot a mistake? Tell us.