Skip to content
Book Open access

Spatiotemporal Load Balancing for Near-Memory Accelerated Databases by Partial Resharding

Sep 2026 · Workshop Proceedings of the 55th International Conference on Parallel Processing · pp. 75-82 · 1 citation · 9 references

TL;DR

This paper proposes an extension of query density-driven partitioning to support dynamically changing workloads and achieves significantly higher throughput than PIM-tree, a skew-resistant state-of-the-art data structure, during periods without workload changes.

Abstract

Near-memory acceleration, where a large number of compute nodes with limited memory process data in parallel, is a promising approach for in-memory databases. Therein, partitioning is required to balance data and queries for skewed workloads. However, existing load balancing methods lack efficient support for dynamic workload changes. As an instance, query density-driven partitioning statically refers to the reference workload to find hot ranges of data and distributes them over less busy nodes. In this paper, we propose an extension of query density-driven partitioning to support dynamically changing workloads. We distribute hot ranges to a limited number of nodes as long as theoretical worst-case load balance is maintained, reserving less busy nodes for hot ranges to appear in future. This approach enables us to restore load balancing with small amortized overhead for dynamic workloads. Our experiments have confirmed that our approach can handle more frequent changes in the workload than query density-driven partitioning. We have also achieved significantly higher throughput with our approach than PIM-tree, a skew-resistant state-of-the-art data structure, during periods without workload changes.

Read PDF

Similar papers

Book Open access Sep 2026

MEDO: Adaptive Multi-Stream Data Offloading for Efficient Management of Disaggregated Memory Pool

MEDO leverages a novel multi-stream data offloading architecture, featuring parallel data streams and approximate LRU queues, to maximize throughput and efficiently handle diverse workloads and incorporates a lightweight, adaptive offloading agent that dynamically optimizes data placement decisions and fine-grained sys...

Jing Wang, Han-Zhang Yang, Chao Li et al. · 0 citations
Jul 2026

CrocSort: Resource-Efficient, Skew-Resilient Parallel External Merge Sort

CrocSort is presented, a byte-balanced parallel external merge sort with configurable memory and per-phase thread settings with practical resource-configuration rules for selecting these settings from input size, memory budget, and thread cap.

Riki Otaki, Charles Benello, Fuheng Zhao et al. · 0 citations
Book Open access Sep 2026

BASIC-Prefetcher: Bin-based Address and Size-Informed Caching for AI-Driven SSD Workloads

The BASIC Prefetcher is proposed, which recovers application-level I/O patterns by classifying requests into size-based bins and modeling temporal bin transitions with a lightweight Markov-chain framework and achieves up to 89% prediction accuracy on sequential workloads and 70% on highly random workloads, including AI...

Han Jang, Dongjun Lee, Youngbin Jin et al. · 0 citations

S !"#$ : A Scalable and Resize-optimized Hash Index on Disaggregated Memory

A novel architecture called S !"#$, designed to enhance the performance of hash indexes in disaggregated memory, is introduced and the results show that S !"#$ outperforms state-of-the-art DM-optimized hash indexes by at most 6.7 → (RACE), 3.6 → (SepHash), and 1.8 → (Outback) in YCSB workloads, respectively.

Han-Tian Zha, Teng Ma, Bao-Tong Lu et al. · 0 citations
Book Open access Aug 2026

STORM: Enabling Traffic Scheduling for RDMA

STORM is presented, a NIC-level scheduler for all types of RDMA workloads using NIC-only information: the known RDMA request size, and per-queue-pair backlog, and converts these signals into a small number of extra priority levels on the wire and prioritizes requests that are either near completion or blocking queued d...

Jichun Wu, Ran Shu, Gianni Antichi et al. · 0 citations

T 𝒆𝒙𝑩𝒆𝒏𝒄𝒉 : A Unified Benchmarking Suite for Shifting Workloads

A unified key-value benchmarking suite that enables benchmarking key-value stores against dynamically shifting and production-like workloads and comparing their performance side by side and allows users to benchmark multiple databases and perform an apples-to-apples comparison under the same workload readily within a s...

Abhishek Chanda, Shubham Kaushik, A. Lavrov et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.