Quicksand: Harnessing Stranded Datacenter Resources with Granular Computing
Zhenyuan Ruan, Shihang Li, Kaiyan Fan, Seo Jin Park, Marcos K. Aguilera, Adam Belay, Malte Schwarzkopf
Abstract
Datacenters today waste CPU and memory, as resources demanded by applications often fail to match the resources available on machines. This leads to stranded resources because one resource that runs out prevents placing additional applications that could consume the other resources. Unusable stranded resources result in reduced utilization of servers, and wasted money and energy.
Quicksand is a new framework and runtime system that unstrands resources by providing developers with familiar, high-level abstractions (e.g., data structures, batch computing). Internally Quicksand decomposes them into resource proclets, granular units that each primarily consume resources of one type. Inspired by recent granular programming models, Quicksand decouples consumption of resources as much as possible. It splits, merges, and migrates resource proclets in milliseconds, so it can use resources on any machine, even if available only briefly.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 98a20c2f-d764-4ea8-9a83-ab3a798c9a9eCited by top-tier papers3
- Burst Computing: Quick, Sudden, Massively Parallel Processing on Serverless ResourcesDaniel Barcelona Pons, Aitor Arjona, Pedro García López, Enrique Molina-Giménez et al.USENIX ATC 2025 · 3 citations
- DDB: Source-Level Interactive Debugging for Distributed ApplicationsYibo Yan, Junzhou He, Seo Jin ParkSOSP 2026
- Continuation-Centric Computing with ArcaAkshay Srivatsan, Yuhan Deng, Katherine Mohr, Emma Sudo et al.OSDI 2026
Builds on13
- Pond: CXL-Based Memory Pooling Systems for Cloud PlatformsHuaicheng Li, Daniel S. Berger, Lisa Hsu, Daniel Ernst et al.ASPLOS 2023 · 328 citations
- Borg: the next generationMuhammad Tirmazi, Adam Barker, Nan Deng, Md E. Haque et al.EuroSys 2020 · 323 citations
- Can far memory improve job throughput?Emmanuel Amaro, Christopher Branner-Augmon, Zhihong Luo, Amy Ousterhout et al.EuroSys 2020 · 163 citations
- Analyzing and Mitigating Data Stalls in DNN TrainingJayashree Mohan, Amar Phanishayee, Ashish Raniwala, Vijay ChidambaramVLDB 2021 · 142 citations
- Beware of Fragmentation: Scheduling GPU-Sharing Workloads with Fragmentation Gradient DescentQizhen Weng, Lingyun Yang, Yinghao Yu, Wei Wang et al.USENIX ATC 2023 · 115 citations
Related papers
- Understanding the Effect of Data Center Resource Disaggregation on Production DBMSsQizhen Zhang, Yifan Cai, Xinyi Chen, Sebastian Angel et al.VLDB 2020 · 64 citations
- Memory deduplication for serverless computing with MedesDivyanshu Saxena, Tao Ji, Arjun Singhvi, Junaid Khalid et al.EuroSys 2022 · 54 citations
- Yield Not Thy CoreAchilles Benetopoulos, Peter Alvaro, Andi Quinn, Robert SouléEuroSys 2026
- RunTime-assisted convergence in replicated data typesGowtham Kaki, Prasanth Prahladan, Nicholas V. LewchenkoPLDI 2022 · 3 citations
- Accessible near-storage computing with FPGAsRobert Schmid, Max Plauth, Lukas Wenzel, Felix Eberhardt et al.EuroSys 2020 · 26 citations
