NEMO is presented, a nimble and expressive hardware memory telemetry engine for server memory controllers (MCs) that gives OS subsystems policy-specific views of memory behavior and provides higher-fidelity signals at substantially lower CPU overhead across a range of state-of-the-art memory management systems.
CAISA is introduced, a composable AI systems architecture that enables disaggregated memory expansion for large-scale AI workloads using CXL-based shared memory and coupling memory isolation with software-managed data orchestration decouples compute and memory resources while preserving efficient data movement, providi...
Divya Kiran Kadiyala, Lianjie Cao, Jinsun Yoo et al.· Conference on Applications,...· 0 citations
CXL-enabled memory expands server memory capacity, but introduces a page-placement problem: the operating system must decide which pages should reside in DRAM and which should reside on slower CXL memory. Existing systems make this tradeoff in one of two ways. Userspace controllers support flexible policies, but expose...
Embodied LLM systems increasingly co-locate latency-critical robotics pipelines with compute- and memory-intensive language-model inference on edge platforms such as NVIDIA Jetson. This co-location avoids cloud round trips and enables privacy-preserving, low-latency interaction, but it also creates a new source of reso...
Teng Mei, Cheng-Xuan Pei, Marco Canini et al.· Proceedings of the 17th ACM...· 1 citation
Dorado is a novel design that scales SmartNIC session tables entirely on inexpensive DDR modules and uses three new techniques that extract commodity DDR performance by restructuring session table layout, decomposing processing pipelines to reduce locking, and scheduling memory accesses to minimize stalls.
Heng Yu, Kai Ren, Jiajun Liang et al.· Conference on Applications,...· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.