Skip to content
Open access

Transfer-Efficient Data Processing in Disaggregated Systems

Aug 2026 · Datenbank-Spektrum · 0 citations · 9 references

TL;DR

Staged query execution approach that interleaves query evaluation with fine-grained remote loading using random-access storage layouts and intermediate selection vectors to fetch only the necessary data for further processing is proposed, showing that fine-grained staged loading significantly reduces transferred data and improves runtime for selective analytical queries.

Abstract

Disaggregated analytics systems separate compute, memory, and storage to improve elasticity and resource utilisation, making network transfer a central bottleneck. Existing systems typically access remote data at fixed coarse granularities, transferring entire files, columns, or column-chunks even when queries process only small subsets of values. This mismatch between query selectivity and transfer granularity creates substantial unnecessary transfer. To address this problem, we propose making remote access granularity a first-class optimisation concern in disaggregated analytical systems. We propose a staged query execution approach that interleaves query evaluation with fine-grained remote loading using random-access storage layouts and intermediate selection vectors to fetch only the necessary data for further processing. By leveraging ideas of composable systems, the approach preserves compatibility with existing vectorised in-memory execution kernels. Using a prototype integrated with the BOSS composable DBMS, we show that fine-grained staged loading significantly reduces transferred data and improves runtime for selective analytical queries. We also show that memoisation allows lazy-loading overhead to converge to zero as an increasing fraction of data is accessed on the compute node. These results suggest that future disaggregated analytical systems should dynamically adapt remote access granularity during query execution to achieve transfer-efficient analytics.

Read PDF

Similar papers

Book Open access Sep 2026

MEDO: Adaptive Multi-Stream Data Offloading for Efficient Management of Disaggregated Memory Pool

MEDO leverages a novel multi-stream data offloading architecture, featuring parallel data streams and approximate LRU queues, to maximize throughput and efficiently handle diverse workloads and incorporates a lightweight, adaptive offloading agent that dynamically optimizes data placement decisions and fine-grained sys...

Jing Wang, Han-Zhang Yang, Chao Li et al. · 0 citations

S !"#$ : A Scalable and Resize-optimized Hash Index on Disaggregated Memory

A novel architecture called S !"#$, designed to enhance the performance of hash indexes in disaggregated memory, is introduced and the results show that S !"#$ outperforms state-of-the-art DM-optimized hash indexes by at most 6.7 → (RACE), 3.6 → (SepHash), and 1.8 → (Outback) in YCSB workloads, respectively.

Han-Tian Zha, Teng Ma, Bao-Tong Lu et al. · 0 citations
Jul 2026

CrocSort: Resource-Efficient, Skew-Resilient Parallel External Merge Sort

CrocSort is presented, a byte-balanced parallel external merge sort with configurable memory and per-phase thread settings with practical resource-configuration rules for selecting these settings from input size, memory budget, and thread cap.

Riki Otaki, Charles Benello, Fuheng Zhao et al. · 0 citations
Book Open access Sep 2026

MiTDM: Eliminating False Conflicts in Scalable MVCC in Disaggregated Memory

MiTDM introduces a hierarchical block version chain that combines the benefits of array and chained structures, supporting dynamic version expansion while maintaining low-latency access, and achieves up to 80.1% higher throughput compared to FaRMV2 and 42.7% higher than Motor.

Ao-Xin Wei, Jin-Tian Wu, Jian Zhou et al. · 0 citations
Preprint Sep 2026

Decoupling Disaggregated Memory Optimizations from Indexing: A Compiler-Runtime Approach

Disaggregated memory (DM) decouples compute and memory into independently scalable pools, connected over a slower interconnect rather than a local bus. This decoupling is exactly what makes DM attractive--but it also means that every index must now reason explicitly about remote-memory access and its associated optimiz...

Xin-Peng Zhao, Ze-Ling Long, Chaichon Wongkham et al. · 0 citations
Jul 2026

The Data World is Not Flat: Efficient Factorized Execution for Relational Systems

A novel code-generating engine with factorization that enables intra-query-parallelized query execution on factorized representations and generates code to overcome their CPU-unfriendly layout, offering a unified and scalable solution for modern workloads.

Stefan Lehner, Thomas Neumann · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.