XBRIDGE is proposed, a decode-free communication protocol that outperforms text-based communication on all seven tasks for each model pair while achieving 11x lower latency, and in a same-architecture setting it also exceeds a KV-sharing baseline on six of seven tasks.
Abstract
Heterogeneous multi-agent LLM systems, where agents are powered by different model families, can outperform homogeneous configurations by reducing redundant reasoning patterns. Yet existing communication protocols either operate through text, discarding the sender's internal representations, or require architectural homogeneity for latent-level transfer. We identify the entity grounding problem in cross-architecture communication: cross-attention bridges that transfer continuous representations across different LLM families suffer from rare-token compression collapse, where entity identity is lost in the continuous bottleneck (bridge-only F1 ~30%). We propose XBRIDGE, a decode-free communication protocol that addresses this through two mechanisms. Lexical Anchor Mapping (LAM) maps the sender's original context tokens to the receiver's vocabulary, providing discrete entity anchors. A Latent Enrichment Bridge (LEB) lets the receiver query the sender's hidden states for contextual enrichment. The entity anchors ground the bridge's contextual signals to specific entities through the receiver's own self-attention. Across three model families (Llama, Qwen, and Mistral), seven benchmarks, and both communication directions, XBRIDGE outperforms text-based communication on all seven tasks for each model pair while achieving 11x lower latency, and in a same-architecture setting it also exceeds a KV-sharing baseline on six of seven tasks. LEB requires only 264M trainable parameters (3.8% of the receiver), is trained on a small balanced sample set, and adds negligible inference overhead.
Multi-agent LLM systems split work across models, so answering often requires knowledge that sits in another agent's context: a Sharer has encoded information that a Receiver needs to complete its task. They usually communicate by exchanging text, which puts autoregressive decoding on the critical path and reduces the...
Ji-Yao Liu, Qi Zhang, Yaoyi Jia et al.· 1 citation
LLM-based multi-agent systems (MAS) increasingly use latent collaboration to avoid the information loss and repeated encoding-decoding overhead of natural-language communication. However, directly forwarding all sender latents makes the receiver-side context scale with both the number of agents and the reasoning length...
Shi-Nan Zhang, Tao Zhang, Qi-Hui Zhu et al.· 0 citations
Results show that HeteroFold enables efficient cross-family KV reuse without receiver prefill, and achieves the best cache-transfer performance on all four long-context benchmarks and most short-context settings.
Vincent-Daniel Yun, Woo-Sang Lim, Haneul Yoo et al.· 0 citations
This work proposes StateBridge, a training-free latent communication approach that aligns the sender's final-layer hidden states to the receiver's input space via a closed-form orthogonal transformation and prepended to the input of the receiver agent as a continuous prefix.
Yan-Wen Peng, Delvin Ce Zhang, Xi Wang et al.· 6 citations
CacheBack is a simple, robust, training-free instance of receiver conditioning based on the sender's attention weights that improves accuracy and reduces median task-completion latency on FanOutQA and shows comparable improvements across model families that span dense Transformers, Mamba-attention hybrids, and sliding-...
M. Rossi, Prajwal Raghunath, Hao Xuan et al.· 0 citations
This work introduces Machine-Interpretable Information (MII), the first agent-to-agent (A2A) document-to-state protocol, and demonstrates strong cross-model interoperability across heterogeneous LLMs -- despite the Writer using a legacy GPT-2 vocabulary, forcing genuine semantic translation rather than token-level memo...
Yi-Fan Wang, De-Jing Dou· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.