The LogDrive: Composable Durability for Cloud-Based Shared Logs
Gardner Vickers, Lucas Bradstreet, Mahesh Balakrishnan, Prince Mahajan, David Mao, Xavier Léauté, Ismael Juma, Nikhil Bhatia, Jack Vanlightly, Prateek Jindal, Sumit Arrawatia, Andrew Grant
摘要
A growing class of systems leverages inexpensive cloud object storage as a disaggregated data plane, offloading durability and scaling to the cloud. Storing the metadata for such systems in a cloud database is too expensive; while self-managed databases are complex and fragile. Conflux provides a third option by storing a shared log on cloud storage and using it to replicate state across VMs. A key innovation in Conflux is the separation of durability from sequencing. Durability is provided solely by the novel LogDrive abstraction: a simple, low-level substrate that can be layered above arbitrary cloud storage, striped for throughput, and – unlike shared logs – composed via quorum-based replication. Sequencing is provided by the AtomicLog, which implements a conventional shared log over any LogDrive. This design allows Conflux to replicate arbitrary state machines while using RAID-like compositions of cloud storage for durability. Conflux is deployed in production at Confluent as the metadata service for an S3-based publish-subscribe system called K2. We show that Conflux can run on diverse storage services (e.g., DynamoDB, S3, S3Express) with just a few hundred extra LOC per service; as well as augment the durability of these services (e.g., with synchronous cross-region replication). Conflux unlocks new cost vs. latency trade-offs over cloud storage: for our representative workloads and latency SLAs (compared to using DynamoDB directly) Conflux-over-DynamoDB slashes metadata cost by 10X and overall cost by 3X.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper9
- Boki: Stateful Serverless Computing with Shared LogsZhipeng Jia, Emmett WitchelSOSP 2021 · 被引用 81 次
- Scalog: Seamless Reconfiguration and Total Order in a Scalable Shared LogCong Ding, David Chu, Evan Zhao, Xiang Li 等NSDI 2020 · 被引用 52 次
- Virtual Consensus in DelosMahesh Balakrishnan, Jason Flinn, Chen Shen, Mihir Dharamshi 等OSDI 2020 · 被引用 42 次
- Gryff: Unifying Consensus and Shared RegistersMatthew Burke, Audrey Cheng, Wyatt LloydNSDI 2020 · 被引用 29 次
- Halfmoon: Log-Optimal Fault-Tolerant Stateful Serverless ComputingSheng Qi, Xuanzhe Liu, Xin JinSOSP 2023 · 被引用 16 次
相关 Paper
- BtrLog: Low-Latency Logging for Cloud Database SystemsMaximilian Kuschewski, Lam-Duy Nguyen, Matthias Jasny, Tobias Ziegler 等VLDB 2026
- Cornus: Atomic Commit for a Cloud DBMS with Storage DisaggregationZhihan Guo, Xinyu Zeng, Kan Wu, Wuh-Chwen Hwang 等VLDB 2023 · 被引用 23 次
- Marlin: Efficient Coordination for Autoscaling Cloud DBMSWenjie Hu, Guanzhou Hu, Mahesh Balakrishnan, Xiangyao YuSIGMOD 2026 · 被引用 1 次
- SWARM: Replicating Shared Disaggregated-Memory Data in No TimeAntoine Murat, Clément Burgelin, Athanasios Xygkis, Igor Zablotchi 等SOSP 2024 · 被引用 2 次
- LoLKV: The Logless, Linearizable, RDMA-based Key-Value Storage SystemAhmed Alquraan, Sreeharsha Udayashankar, Virendra J. Marathe, Bernard Wong 等NSDI 2024 · 被引用 5 次
