This work presents file-backed weight adoption: a framework-independent producer maps each tensor with MAP_SHARED, wraps the pages as a no-copy GPU buffer, and exports a DLPack capsule that PyTorch or MLX imports as ordinary storage.
Yuan Si, Yufeng Lin, Da-Ming Li et al.· 1 citation
This work instantiates budgeted oracle-to-hint compression in online judge (OJ) style algorithmic programming as a modular interactive agent that couples an LLM core with a sandboxed judger, a feedback-to-hint prompt constructor, and trajectory memory.
Jialiang Gu, Keren Zhou, Daming Li et al.· SIGSOFT FSE Companion· 2 citations
The resulting design principle is simple: in this regime, let the kernel own eviction, while model-specific knowledge is best spent on admission and advice.
Yuan Si, Yufeng Lin, Daming Li et al.· 2 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.