Skip to content

Seron: Smart Query Router for Multi-Primary Cloud-Native Databases with Shared Storage

Aug 2026 · Proceedings of the VLDB Endowment · Vol 19, pp. 4195-4208 · 0 citations · 51 references

TL;DR

Seron is presented, a smart query routing system customized for multi-primary cloud-native databases with shared storage that generates workload-aware logical routing rules at different granularities using a two-stage framework, and actively detects and handles physical data collisions through judicious conflict modeling.

Abstract

Cloud-native databases have become a major deployment paradigm for commercial applications and services in recent years. One of the most promising development trends in cloud-native databases is to support multi-primary architecture with shared storage layers. However, when different writer nodes simultaneously access overlapping data ranges, query conflicts naturally arise, potentially leading to a significant performance drop. Therefore, it is of great importance to design query routers that can smartly separate incoming workloads to mitigate inter-node conflicts, which is a problem not fully discussed in the existing literature. To address this issue, in this paper, we present Seron, a smart query routing system customized for multi-primary cloud-native databases with shared storage. The system aims to minimize conflicts by directing each query to an optimal compute node. To this end, Seron generates workload-aware logical routing rules at different granularities using a two-stage framework, actively detects and handles physical data collisions through judicious conflict modeling, and constantly adapts its routing policies to provide the best runtime performance. The system has been integrated into GaussDB and deployed in real-world scenarios. Extensive experiments verify the effectiveness and cost-efficiency of Seron, which achieves up to 61.7% execution time reduction compared to the best baseline.

View source

Similar papers

Towards Practical Latency SLOs on Cloud Data Warehouses

This work outlines AutoSLO, a latency-SLO-aware work-load management framework for multi-cluster cloud data ware-houses that includes a periodic Policy Tuner that proactively plans resources using workload fore-casts, an SLO-aware reactive Autoscaler that adjusts the active cluster set based on the observed workload, a...

Markos Markakis, T. Kraska · 0 citations
Open access Sep 2026

CacheServer: Disaggregated Caching for Cloud Databases

Cloud databases typically cache data fetched from cloud storage to compute nodes and conduct affinity scheduling to guarantee data locality -- scheduling queries accessing the same data segment to the same node. Although this strategy improves the cache hit rate and effectively reduces the prohibitive data transmis...

Jun-Yong Zhao, Jia Yuan, Lei Cao · 0 citations
Open access Sep 2026

EASE: Resource-aware Query Scheduling across Heterogeneous Cloud Compute Services

Modern cloud platforms provide diverse compute services, including virtual machines, Function-as-a-Service, and Query-as-a-Service, each offering unique trade-offs in performance, elasticity, and cost. While these services collectively cover the diverse needs of OLAP workloads, existing systems typically rely on a sing...

Wen-Bo Li, Hao-Qiong Bian, Chao Zhang et al. · 0 citations
Open access May 2026

A Resource-centric Analysis and Optimization of NoSQL Workloads using Distressed Resource Volume Metric

This work proposes and develops an open-source policy simulation framework, LoadStar, which forms a reusable benchmark pipeline for validating policies for resource-centric NoSQL workloads, and defines a resource optimization problem for placing Cosmos DB replicas onto VM nodes, and develops the Luna model for forecast...

Gunika Verma, V. AashutoshA, P. Srinivas et al. · 0 citations
Open access Aug 2026

Five-Minute Rule for Data Analytics in the Cloud

Analysis on AWS shows that caches are beneficial when a system makes two requests per hour for latency-sensitive workloads, or seven requests per second for non-latency-sensitive workloads, which is consistent with and helps explain the near ubiquity of object store caches in cloud analytics systems.

Kira Duwe, Andrew Lamb, L. Lersch et al. · 0 citations

S !"#$ : A Scalable and Resize-optimized Hash Index on Disaggregated Memory

A novel architecture called S !"#$, designed to enhance the performance of hash indexes in disaggregated memory, is introduced and the results show that S !"#$ outperforms state-of-the-art DM-optimized hash indexes by at most 6.7 → (RACE), 3.6 → (SepHash), and 1.8 → (Outback) in YCSB workloads, respectively.

Han-Tian Zha, Teng Ma, Bao-Tong Lu et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.