GA
Gustavo Alonso
cs.DBcs.ARcs.DCcs.LGcs.AIcs.IRcs.NIcs.CLcs.OScs.DS
On Valency
published · living versionsW_u7fa5rax·v1 · currentpublished
MLSys: The New Frontier of Machine Learning Systems
with Alexander Ratner, Dan Alistarh, David G. Andersen, Peter Bailis +64
1 version
Preprints & journals
46 papers in the corpus · 2012–2026Over the Memory Wall, Into the Instruction Wall: The New Bottleneck in GPU Data Processing2608.13696v1 · Sven Hepkema, Bowen Wu, Christos Kozyrakis et al.2026 · 0 citationsarXiv
Oasis: Hiding the Cost of Querying Parquet Files in the Datapath2608.02268v1 · Jonas Dann, Luca Tagliavini, Gustavo Alonso2026 · 0 citationsarXiv
Eiger: An Efficient Library for GPU-based Data Analytics2607.04489v1 · Bowen Wu, Marko Kabi'c, Sven Hepkema et al.2026 · 0 citationsarXiv
Approaching Shannon Bound with Lossless LLM Weight Compression2606.15789v1 · Hongshi Tan, Yao Chen, Gustavo Alonso et al.2026 · 0 citationsarXiv
ObjectCache: Layerwise Object-Storage Retrieval for KV Cache Reuse2605.22850v1 · Yu Zhu, Aditya Dhakal, Yunming Xiao et al.2026 · 0 citationsarXiv
To GPU or Not to GPU: Vector Search in Relational Engines2605.15957v1 · Vasilis Mageirakos, Joel Andr'e, Marko Kabi'c et al.2026 · 0 citationsarXiv
Should I Hide My Duck in the Lake?2602.18775v2 · Jonas Dann, Gustavo Alonso2026 · 0 citationsarXiv
SCENIC: Stream Computation-Enhanced SmartNIC2604.15128v1 · Benjamin Ramhorst, Maximilian Jakob Heer, Luhao Liu et al.2026 · 1 citationarXiv
RoCE BALBOA: Service-enhanced Data Center RDMA for SmartNICs2507.20412v1 · Maximilian Jakob Heer, Benjamin Ramhorst, Yu Zhu et al.2025 · 1 citationarXiv
Fast Graph Vector Search via Hardware Acceleration and Delayed-Synchronization Traversal2406.12385v2 · Wenqi Jiang, Hang Hu, Torsten Hoefler et al.2024 · 7 citationsProceedings of the VLDB Endowment Volume 18, 2025
WARP: An Efficient Engine for Multi-Vector Retrieval2501.17788v3 · Jan Luca Scheerer, Matei Zaharia, Christopher Potts et al.2025 · 8 citationsarXiv
Coyote v2: Raising the Level of Abstraction for Data Center FPGAs2504.21538v1 · Benjamin Ramhorst, Dario Korolija, Maximilian Jakob Heer et al.2025 · 9 citationsarXiv
The Cambridge Report on Database Research2504.11259v1 · Anastasia Ailamaki, Samuel Madden, Daniel Abadi et al.2025 · 3 citationsarXiv
Chameleon: a Heterogeneous and Disaggregated Accelerator System for Retrieval-Augmented Language Models2310.09949v4 · Wenqi Jiang, Marco Zeller, Roger Waleffe et al.2023 · 17 citationsarXiv
SwiftSpatial: Spatial Joins on Modern Hardware2309.16520v2 · Wenqi Jiang, Oleh-Yevhen Khavrona, Martin Parvanov et al.2023 · 3 citationsarXiv
RAGO: Systematic Performance Optimization for Retrieval-Augmented Generation Serving2503.14649v2 · Wenqi Jiang, Suvinay Subramanian, Cat Graves et al.2025 · 16 citationsarXiv
FpgaHub: Fpga-centric Hyper-heterogeneous Computing Platform for Big Data Analytics2503.09318v1 · Zeke Wang, Jie Zhang, Hongjing Huang et al.2025 · 1 citationarXiv
Efficiently Processing Joins and Grouped Aggregations on GPUs2312.00720v2 · Bowen Wu, Dimitrios Koutsoukos, Gustavo Alonso2023 · 15 citations2025. Proc. ACM Manag. Data 3, 1, Article 39 (February 2025), 27 pages
Efficient Tabular Data Preprocessing of ML Pipelines2409.14912v1 · Yu Zhu, Wenqi Jiang, Gustavo Alonso2024 · 1 citationarXiv
CXL and the Return of Scale-Up Database Engines2401.01150v2 · Alberto Lerner, Gustavo Alonso2024 · 27 citationsarXiv
Boxer: FaaSt Ephemeral Elasticity for Off-the-Shelf Cloud Applications2407.00832v1 · Michael Wawrzoniak, Rodrigo Bruno, Ana Klimovic et al.2024 · 2 citationsarXiv
Imaginary Machines: A Serverless Model for Cloud Applications2407.00839v1 · Michael Wawrzoniak, Rodrigo Bruno, Ana Klimovic et al.2024 · 0 citationsarXiv
TablePuppet: A Generic Framework for Relational Federated Learning2403.15839v1 · Lijie Xu, Chulin Xie, Yiran Guo et al.2024 · 0 citationsarXiv
Demystifying Graph Databases: Analysis and Taxonomy of Data Organization, System Designs, and Graph Queries1910.09017v8 · Maciej Besta, Robert Gerstenberger, Emanuel Peter et al.2019 · 146 citationsarXiv
Co-design Hardware and Algorithm for Vector Search2306.11182v3 · Wenqi Jiang, Shigang Li, Yu Zhu et al.2023 · 21 citationsarXiv
Data Processing with FPGAs on Modern Architectures2304.03044v3 · Wenqi Jiang, Dario Korolija, Gustavo Alonso2023 · 6 citationsarXiv
Resource Allocation in Serverless Query Processing2208.09519v1 · Simon Kassing, Ingo Muller, Gustavo Alonso2022 · 3 citationsarXiv
ECI: a Customizable Cache Coherency Stack for Hybrid FPGA-CPU Architectures2208.07124v1 · Abishek Ramdas, Michael Giardino, Runbin Shi et al.2022 · 1 citationarXiv
How to use Persistent Memory in your Database2112.00425v1 · Dimitrios Koutsoukos, Raghav Bhartia, Ana Klimovic et al.2021 · 0 citationsarXiv
Evaluating Query Languages and Systems for High-Energy Physics Data [Extended Version]2104.12615v3 · Dan Graur, Ingo Muller, Mason Proffitt et al.2021 · 14 citationsarXiv
Modularis: Modular Relational Analytics over Heterogeneous Distributed Platforms2004.03488v2 · Dimitrios Koutsoukos, Ingo Muller, Renato Marroqu'in et al.2020 · 1 citationarXiv
From Research to Proof-of-Concept: Analysis of a Deployment of FPGAs on a Commercial Search Engine2108.09073v1 · Fabio Maschi, Gustavo Alonso, Anthony Hock-Koon et al.2021 · 1 citationarXiv
Farview: Disaggregated Memory with Operator Off-loading for Database Engines2106.07102v1 · Dario Korolija, Dimitrios Koutsoukos, Kimberly Keeton et al.2021 · 6 citationsarXiv
MicroRec: Efficient Recommendation Inference by Hardware and Data Structure Solutions2010.05894v2 · Wenqi Jiang, Zhenhao He, Shuai Zhang et al.2020 · 11 citationsarXiv
Rumble: Data Independence for Large Messy Data Sets1910.11582v3 · Ingo Muller, Ghislain Fourny, Stefan Irimescu et al.2019 · 2 citationsarXiv
HyperLogLog Sketch Acceleration on FPGA2005.13332v2 · Amit Kulkarni, Monica Chiosa, Thomas B. Preu\sser et al.2020 · 18 citationsarXiv
Benchmarking High Bandwidth Memory on FPGAs2005.04324v1 · Zeke Wang, Hongjing Huang, Jie Zhang et al.2020 · 5 citationsarXiv
Lambada: Interactive Data Analytics on Cold Data using Serverless Cloud Infrastructure1912.00937v1 · Ingo Muller, Renato Marroqu'in, Gustavo Alonso2019 · 136 citationsarXiv
Using DSP Slices as Content-Addressable Update Queues2004.11080v1 · Thomas B. Preu\sser, Monica Chiosa, Alexander Weiss et al.2020 · 3 citationsarXiv
The Collection Virtual Machine: An Abstraction for Multi-Frontend Multi-Backend Data Analysis2004.01908v2 · Ingo Muller, Renato Marroqu'in, Dimitrios Koutsoukos et al.2020 · 2 citationsarXiv
Reproducible Floating-Point Aggregation in RDBMSs1802.09883v1 · Ingo Muller, Andrea Arteaga, Torsten Hoefler et al.2018 · 8 citations34th IEEE International Conference on Data Engineering (ICDE) 2018, Paris, 2018, pp. 1049-1060
Pay One, Get Hundreds for Free: Reducing Cloud Costs through Shared Query Execution1809.00159v1 · Renato Marroqu'in, Ingo Muller, Darko Makreshanski et al.2018 · 13 citationsProceedings of the ACM Symposium on Cloud Computing (SoCC) 2018, pages 439-450
High Bandwidth Memory on FPGAs: A Data Analytics Perspective2004.01635v1 · Kaan Kara, Christoph Hagleitner, Dionysios Diamantopoulos et al.2020 · 32 citationsarXiv
MLSys: The New Frontier of Machine Learning Systems1904.03257v3 · Alexander Ratner, Dan Alistarh, Gustavo Alonso et al.2019 · 20 citationsarXivon Valency
Accelerating Generalized Linear Models with MLWeaving: A One-Size-Fits-All System for Any-precision Learning (Technical Report)1903.03404v2 · Zeke Wang, Kaan Kara, Hantian Zhang et al.2019 · 4 citationsPVLDB, 2019
SharedDB: Killing One Thousand Queries With One Stone1203.0056v1 · Georgios Giannikis, Gustavo Alonso, Donald Kossmann2012 · 16 citationsProceedings of the VLDB Endowment (PVLDB), Vol. 5, No. 6, pp. 526-537 (2012)
Career total: 430 works. 46 are in this corpus.
Profile built from the corpus for this byline.
Author records are still filling in while the Hub is in alpha. If this is your page, you'll be able to claim it soon. Spot a mistake? Tell us.