cs.LGstat.MLcs.AIcs.CVcs.CLbioinformaticscs.DCstat.MEModels, Geneticq-bio.QM

On Valency

published · living versions
W_u7fa5rax·v1 · currentpublished
MLSys: The New Frontier of Machine Learning Systems
with Alexander Ratner, Dan Alistarh, Gustavo Alonso, David G. Andersen +64
1 version

Preprints & journals

445 papers in the corpus · 1999–2026
How to Build the Virtual Cell with Artificial Intelligence: Priorities and Opportunities.39398201 · Bunne, Charlotte, Roohani, Yusuf, Rosen, Yanay et al.2026 · 385 citationsArXiv. 2024
Esoteric Language Models: A Family of Any-Order Diffusion LLMs2506.01928v5 · Subham Sekhar Sahoo, Zhihan Yang, Yash Akhauri et al.2025 · 0 citationsarXiv
PertAdapt: unlocking single-cell foundation models for genetic perturbation prediction via condition-sensitive adaptation.42412811 · Bai, Ding, Song, Le, Xing, Eric P2026 · 0 citationsBioinformatics (Oxford, England). 2026;42(1)
Petuum: A New Platform for Distributed Machine Learning on Big Data1312.7651v2 · Eric P. Xing, Qirong Ho, Wei Dai et al.2013 · 407 citationsarXiv
General Agentic Planning Through Simulative Reasoning with World Models2507.23773v3 · Mingkai Deng, Jinyu Hou, Zhiting Hu et al.2025 · 1 citationarXiv
Advancing Reasoning in Diffusion Language Models with Denoising Process Rewards2510.01544v2 · Shaoan Xie, Lingjing Kong, Xiangchen Song et al.2025 · 0 citationsarXiv
SmartCLIP: Modular Vision-language Alignment with Identification Guarantees2507.22264v2 · Shaoan Xie, Lingjing Kong, Yujia Zheng et al.2025 · 0 citationsarXiv
Scaling Structure Aware Virtual Screening to Billions of Molecules with SPRINT.39975427 · McNutt, Andrew T, Adduri, Abhinav K, Ellington, Caleb N et al.2026 · 3 citationsArXiv. 2025
Log-Linear Attention2506.04761v3 · Han Guo, Songlin Yang, Tarushii Goel et al.2025 · 0 citationsarXiv
LAPS: A Length-Aware-Prefill LLM Serving System2601.11589v2 · Jianshu She, Zonghang Li, Hongchao Du et al.2026 · 0 citationsarXiv
In-context Learning of Evolving Data Streams with Tabular Foundational Models2502.16840v2 · Afonso Lourenco, Joao Gama, Eric P. Xing et al.2025 · 1 citationarXiv
Generative AI for Biosciences: Emerging Threats and Roadmap to Biosecurity2510.15975v2 · Zaixi Zhang, Souradip Chakraborty, Amrit Singh Bedi et al.2025 · 0 citationsarXiv
VSA: Faster Video Diffusion with Trainable Sparse Attention2505.13389v5 · Peiyuan Zhang, Yongqi Chen, Haofeng Huang et al.2025 · 0 citationsarXiv
Efficient Long-context Language Model Training by Core Attention Disaggregation2510.18121v1 · Yonghao Zhuang, Junda Chen, Bo Pang et al.2025 · 0 citationsarXiv
Response to Promises and Pitfalls of Deep Kernel Learning2509.21228v1 · Andrew Gordon Wilson, Zhiting Hu, Ruslan Salakhutdinov et al.2025 · 0 citationsarXiv
CITER: Collaborative Inference for Efficient Large Language Model Decoding with Token-Level Routing2502.01976v6 · Wenhao Zheng, Yixiao Chen, Weitong Zhang et al.2025 · 0 citationsarXiv
A call for built-in biosecurity safeguards for generative AI tools.40295784 · Wang, Mengdi, Zhang, Zaixi, Bedi, Amrit Singh et al.2025 · 30 citationsNature biotechnology. 2025;43(6):845-847
Pruning Spurious Subgraphs for Graph Out-of-Distribution Generalization2506.05957v4 · Tianjun Yao, Haoxuan Li, Yongqiang Chen et al.2025 · 0 citationsarXiv
Memory-adaptive Depth-wise Heterogeneous Federated Learning2303.04887v3 · Kai Zhang, Yutong Dai, Hongyi Wang et al.2023 · 1 citationarXiv
Data Mixing Optimization for Supervised Fine-Tuning of Large Language Models2508.11953v1 · Yuan Li, Zhengzhong Liu, Eric Xing2025 · 0 citationsarXiv
Vision-G1: Towards General Vision Language Reasoning with Multi-Domain Data Curation2508.12680v1 · Yuheng Zha, Kun Zhou, Yujia Wu et al.2025 · 0 citationsarXiv
How Does Controllability Emerge In Language Models During Pretraining?2508.01892v1 · Jianshu She, Xinyue Li, Eric Xing et al.2025 · 0 citationsarXiv
Towards Open-World Generation of Stereo Images and Unsupervised Matching2503.12720v2 · Feng Qiao, Zhexiao Xiong, Eric Xing et al.2025 · 1 citationarXiv
Nile-Chat: Egyptian Language Models for Arabic and Latin Scripts2507.04569v1 · Guokan Shang, Hadi Abdine, Ahmad Chamma et al.2025 · 0 citationsarXiv
Uncertainty-Aware Discrete Diffusion Improves Protein Design10.1101/2025.06.30.662407v1 · Mahbub, S., Feinauer, C., Ellington, C. N. et al.2025 · 0 citationsbioRxiv
Global and Local Entailment Learning for Natural World Imagery2506.21476v1 · Srikumar Sastry, Aayush Dhakal, Eric Xing et al.2025 · 0 citationsarXiv
Active learning to classify macromolecular structures in situ for less supervision in cryo-electron tomography.33620460 · Du, Xuefeng, Wang, Haohan, Zhu, Zhenxi et al.2025 · 14 citationsBioinformatics (Oxford, England). 2021;37(16):2340-2346
lmgame-Bench: How Good are LLMs at Playing Games?2505.15146v2 · Lanxiang Hu, Mingjia Huo, Yuxuan Zhang et al.2025 · 0 citationsarXiv
Mixed Membership Stochastic Blockmodels.21701698 · Airoldi, Edoardo M, Blei, David M, Fienberg, Stephen E et al.2025 · 1,506 citationsJournal of machine learning research : JMLR. 2008;9:1981-2014
QuARI: Query Adaptive Retrieval Improvement2505.21647v1 · Eric Xing, Abby Stylianou, Robert Pless et al.2025 · 0 citationsarXiv
ConText-CIR: Learning from Concepts in Text for Composed Image Retrieval2505.20764v1 · Eric Xing, Pranavi Kolouju, Robert Pless et al.2025 · 3 citationsarXiv
Learning to estimate sample-specific transcriptional networks for 7,000 tumors.40408406 · Ellington, Caleb N, Lengerich, Benjamin J, Watkins, Thomas B K et al.2025 · 1 citationProceedings of the National Academy of Sciences of the United States of America. 2025;122(21):e2411930122
A foundation model of transcription across human cell types.39779852 · Fu, Xi, Mo, Shentong, Buendia, Alejandro et al.2025 · 117 citationsNature. 2025;637(8047):965-973
Learning to Estimate Sample-specific Transcriptional Networks for 7000 Tumors10.1101/2023.12.01.569658v2 · Ellington, C. N., Lengerich, B. J., Watkins, T. B. et al.2023 · 6 citationsbioRxiv
Interaction of N-methylmesoporphyrin IX with a hybrid left-/right-handed G-quadruplex motif from the promoter of the SLC2A1 gene.39704129 · Seth, Paul, Xing, Eric, Hendrickson, Andrew D et al.2025 · 6 citationsNucleic acids research. 2025;53(2)
Token Level Routing Inference System for Edge Devices2504.07878v1 · Jianshu She, Wenhao Zheng, Zhengzhong Liu et al.2025 · 1 citationarXiv
RANGE: Retrieval Augmented Neural Fields for Multi-Resolution Geo-Embeddings2502.19781v2 · Aayush Dhakal, Srikumar Sastry, Subash Khanal et al.2025 · 5 citationsarXiv
Causal Representation Learning from Multi-modal Biomedical Observations.40160453 · Sun, Yuewen, Kong, Lingjing, Chen, Guangyi et al.2025 · 0 citationsArXiv. 2025
MegaMath: Pushing the Limits of Open Math Corpora2504.02807v1 · Fan Zhou, Zengzhi Wang, Nikhil Ranjan et al.2025 · 0 citationsarXiv
VideoGLaMM: A Large Multimodal Model for Pixel-Level Visual Grounding in Videos2411.04923v3 · Shehan Munasinghe, Hanan Gani, Wenqi Zhu et al.2024 · 15 citationsarXiv
good4cir: Generating Detailed Synthetic Captions for Composed Image Retrieval2503.17871v1 · Pranavi Kolouju, Eric Xing, Robert Pless et al.2025 · 0 citationsarXiv
Causal Representation Learning from Multimodal Biomedical Observations2411.06518v3 · Yuewen Sun, Lingjing Kong, Guangyi Chen et al.2024 · 1 citationarXiv
Career total: 859 works. 445 are in this corpus.Showing the 50 most recent.

Profile built from the corpus for this byline.

Author records are still filling in while the Hub is in alpha. If this is your page, you'll be able to claim it soon. Spot a mistake? Tell us.