AR
Alexander Ratner
cs.LGstat.MLcs.AIcs.CLcs.CVcs.DBgeneticsstat.APcs.DCcs.MM
On Valency
published · living versionsW_u7fa5rax·v1 · currentpublished
MLSys: The New Frontier of Machine Learning Systems
with Dan Alistarh, Gustavo Alonso, David G. Andersen, Peter Bailis +64
1 version
Preprints & journals
37 papers in the corpus · 2016–2025Towards Acyclic Preference Evaluation of Language Models via Multiple Evaluators2410.12869v5 · Zhengyu Hu, Jieyu Zhang, Zhihan Xiong et al.2024 · 1 citationarXiv
Snorkel: rapid training data creation with weak supervision.32214778 · Ratner, Alexander, Bach, Stephen H, Ehrenberg, Henry et al.2025 · 794 citationsThe VLDB journal : very large data bases : a publication of the VLDB Endowment. 2020;29(2):709-730
Slice-based Learning: A Programming Model for Residual Learning in Critical Data Slices.31871391 · Chen, Vincent S, Wu, Sen, Weng, Zhenzhen et al.2024 · 42 citationsAdvances in neural information processing systems. 2019;32:9392-9402
Found in the Middle: Calibrating Positional Attention Bias Improves Long Context Utilization2406.16008v2 · Cheng-Yu Hsieh, Yung-Sung Chuang, Chun-Liang Li et al.2024 · 27 citationsarXiv
Learning to Compose Domain-Specific Transformations for Data Augmentation.29375240 · Ratner, Alexander J, Ehrenberg, Henry R, Hussain, Zeshan et al.2024 · 264 citationsAdvances in neural information processing systems. 2017;30:3239-3249
MaskSearch: Querying Image Masks at Scale2305.02375v2 · Dong He, Jieyu Zhang, Maureen Daum et al.2023 · 0 citationsarXiv
DataComp: In search of the next generation of multimodal datasets2304.14108v5 · Samir Yitzhak Gadre, Gabriel Ilharco, Alex Fang et al.2023 · 97 citationsarXiv
Learning the Structure of Generative Models without Labeled Data.30882087 · Bach, Stephen H, He, Bryan, Ratner, Alexander et al.2023 · 113 citationsProceedings of machine learning research. 2017;70:273-82
Tool Documentation Enables Zero-Shot Tool-Usage with Large Language Models2308.00675v1 · Cheng-Yu Hsieh, Si-An Chen, Chun-Liang Li et al.2023 · 8 citationsarXiv
Training Complex Models with Multi-Task Weak Supervision.31565535 · Ratner, Alexander, Hancock, Braden, Dunnmon, Jared et al.2023 · 156 citationsProceedings of the ... AAAI Conference on Artificial Intelligence. AAAI Conference on Artificial Intelligence. 2019;33:4763-4771
Nemo: Guiding and Contextualizing Weak Supervision for Interactive Data Programming2203.01382v3 · Cheng-Yu Hsieh, Jieyu Zhang, Alexander Ratner2022 · 4 citationsProceedings of the VLDB Endowment, 15(13): 4093 - 4105, 2022
Leveraging Instance Features for Label Aggregation in Programmatic Weak Supervision2210.02724v2 · Jieyu Zhang, Linxin Song, Alexander Ratner2022 · 3 citationsarXiv
Binary Classification with Positive Labeling Sources2208.01704v1 · Jieyu Zhang, Yujing Wang, Yaming Yang et al.2022 · 4 citationsarXiv
Creating Training Sets via Weak Indirect Supervision2110.03484v3 · Jieyu Zhang, Bohan Wang, Xiangchen Song et al.2021 · 3 citationsarXiv
Cross-Modal Data Programming Enables Rapid Medical Machine Learning.32776018 · Dunnmon, Jared A, Ratner, Alexander J, Saab, Khaled et al.2022 · 60 citationsPatterns (New York, N.Y.). 2020;1(2)
A Survey on Programmatic Weak Supervision2202.05433v2 · Jieyu Zhang, Cheng-Yu Hsieh, Yue Yu et al.2022 · 37 citationsarXiv
WRENCH: A Comprehensive Benchmark for Weak Supervision2109.11377v2 · Jieyu Zhang, Yue Yu, Yinghao Li et al.2021 · 39 citationsarXiv
AMELIE speeds Mendelian diagnosis by matching patient phenotype and genotype to primary literature.32434849 · Birgmeier, Johannes, Haeussler, Maximilian, Deisseroth, Cole A et al.2021 · 131 citationsScience translational medicine. 2020;12(544)
Snorkel DryBell: A Case Study in Deploying Weak Supervision at Industrial Scale.31777414 · Bach, Stephen H, Rodriguez, Daniel, Liu, Yintao et al.2020 · 24 citationsProceedings. ACM-SIGMOD International Conference on Management of Data. 2019;2019:362-375
A Kernel Theory of Modern Data Augmentation.31777848 · Dao, Tri, Gu, Albert, Ratner, Alexander J et al.2020 · 107 citationsProceedings of machine learning research. 2019;97:1528-1537
Slice-based Learning: A Programming Model for Residual Learning in Critical Data Slices1909.06349v2 · Vincent S. Chen, Sen Wu, Zhenzhen Weng et al.2019 · 42 citationsarXiv
A machine-compiled database of genome-wide association studies.31350405 · Kuleshov, Volodymyr, Ding, Jialin, Vo, Christopher et al.2019 · 29 citationsNature communications. 2019;10(1):3341
MLSys: The New Frontier of Machine Learning Systems1904.03257v3 · Alexander Ratner, Dan Alistarh, Gustavo Alonso et al.2019 · 20 citationsarXivon Valency
Data Programming: Creating Large Training Sets, Quickly.29872252 · Ratner, Alexander, De Sa, Christopher, Wu, Sen et al.2019 · 501 citationsAdvances in neural information processing systems. 2016;29:3567-3575
AMELIE 2 speeds up Mendelian diagnosis by matching patient phenotype & genotype to primary literature10.1101/839878v1 · Birgmeier, J., Haeussler, M., Deisseroth, C. A. et al.2019 · 2 citationsbioRxiv
Snorkel DryBell: A Case Study in Deploying Weak Supervision at Industrial Scale1812.00417v2 · Stephen H. Bach, Daniel Rodriguez, Yintao Liu et al.2018 · 24 citationsProceedings of the International Conference on Management of Data (SIGMOD), 2019
Cross-Modal Data Programming Enables Rapid Medical Machine Learning1903.11101v1 · Jared Dunnmon, Alexander Ratner, Nishith Khandwala et al.2019 · 60 citationsarXiv
A Kernel Theory of Modern Data Augmentation1803.06084v2 · Tri Dao, Albert Gu, Alexander J. Ratner et al.2018 · 107 citationsarXiv
Learning Dependency Structures for Weak Supervision Models1903.05844v1 · Paroma Varma, Frederic Sala, Ann He et al.2019 · 36 citationsarXiv
Data Programming: Creating Large Training Sets, Quickly1605.07723v3 · Alexander Ratner, Christopher De Sa, Sen Wu et al.2016 · 501 citationsAdvances in Neural Information Processing Systems 29, 2016, 3567--3575
Learning to Compose Domain-Specific Transformations for Data Augmentation1709.01643v3 · Alexander J. Ratner, Henry R. Ehrenberg, Zeshan Hussain et al.2017 · 264 citationsAdvances in Neural Information Processing Systems 30, 2017, 3236--3246
Training Complex Models with Multi-Task Weak Supervision1810.02840v2 · Alexander Ratner, Braden Hancock, Jared Dunnmon et al.2018 · 156 citationsarXiv
Snorkel: Rapid Training Data Creation with Weak Supervision1711.10160v1 · Alexander Ratner, Stephen H. Bach, Henry Ehrenberg et al.2017 · 795 citationsProceedings of the VLDB Endowment, 11(3), 269-282, 2017
Learning the Structure of Generative Models without Labeled Data1703.00854v2 · Stephen H. Bach, Bryan He, Alexander Ratner et al.2017 · 113 citationsProceedings of the 34th International Conference on Machine Learning, Sydney, Australia, PMLR 70, 2017
AMELIE accelerates Mendelian patient diagnosis directly from the primary literature10.1101/171322v1 · Birgmeier, J., Haeussler, M., Deisseroth, C. A. et al.2017 · 23 citationsbioRxiv
SwellShark: A Generative Model for Biomedical Named Entity Recognition without Labeled Data1704.06360v1 · Jason Fries, Sen Wu, Alex Ratner et al.2017 · 89 citationsarXiv
Career total: 45 works. 37 are in this corpus.Showing the 36 most recent.
Profile built from the corpus for this byline.
Author records are still filling in while the Hub is in alpha. If this is your page, you'll be able to claim it soon. Spot a mistake? Tell us.