Skip to content

Mosaic: Runtime-Efficient Multi-Agent Embodied Planning

Jul 2026 · arXiv.org · Vol abs/2607.09603 · 0 citations · 68 references
Computer Science

TL;DR

Mosaic is introduced, a runtime-efficient multi-agent planning framework that maintains accurate yet lightweight state tracking through agent-centric semantic memory that stores objects in relative coordinates, enabling geometric transformations and coordination.

Abstract

LLM-based multi-agent embodied planning remains impractical due to prohibitively high execution latency. We identify failed actions as the dominant bottleneck, stemming from two core challenges: inaccurate state tracking under partial observability and inefficient coordination that produces redundant or conflicting actions. We introduce Mosaic, a runtime-efficient multi-agent planning framework that addresses both challenges. Mosaic maintains accurate yet lightweight state tracking through agent-centric semantic memory that stores objects in relative coordinates, enabling geometric transformations and coordination. It ensures efficient coordination through Integer Linear Programming that allocates actions at every planning step, enforcing physical feasibility and inter-agent coordination constraints. Across AI2-THOR and search-and-rescue benchmarks, Mosaic achieves 27-32% faster execution, 30-33% fewer LLM calls, 25-31% fewer steps, and 4-10% points higher success rates. These results demonstrate that efficient memory and constraint-guided coordination are critical for scalable, low-latency multi-agent planning.

View source

Similar papers

#artificial intelligence Open access Aug 2026

Generalizable Multi-Agent Planning From Signal Temporal Logic Specifications via Diffusion

A new diffusion method for multi-agent planning with STL specifications is introduced, making the approach generalizable to novel formulas whose predicates are placed anywhere within the goal region covered during training, while achieving the same scalability as existing learning-based methods.

Joe Eappen, Zikang Xiong, S. Iyengar et al. · 0 citations
Preprint Sep 2026

HarnessVLN: Unifying Training-Free Embodied Navigation through an Agent Harness

Embodied navigation requires agents to ground instructions or object goals in spatial observations and translate plans into successful execution. As multimodal large language models (MLLMs) become increasingly capable, they offer stronger support for navigation without task-specific training; however, improved semantic...

Yang Chen, Li-Rong Che, Zhen-Yu Huang et al. · 5 citations · ⚡1
Preprint Aug 2026

Discovering Diverse Planning Policies for Multimodal Embodied Agents with Quality-Diversity Optimization

Multimodal embodied agents are increasingly required to solve long-horizon tasks by integrating visual observations, textual goals, and interaction history into closed-loop decision making. However, state-of-the-art large-model-based planners often rely on a single dominant planning style during execution. Once this ex...

Peng Xu, Yong Liu, Xiaoya Nan et al. · 0 citations
Preprint Aug 2026

CoCoBench: A Cooperative Coordination Benchmark for Embodied Multi-Agent Task Planning

CoBench is introduced, a construct-level benchmark for evaluating multi-agent embodied coordination in executable household tasks and shows that coordination ability is highly construct-specific: strong overall performance does not imply balanced competence across different coordination types.

Yang Chen, Ye-Xin Xie, Li-Rong Che et al. · 0 citations
Preprint Aug 2026

LocalLSTC: A Long Short-Term Control Architecture for Locally Deployed GUI Agents

LocalLSTC is introduced, a training-free architecture that organizes control by temporal scope, maintaining persistent cross-step state to guide short-term execution commitments, and identifies temporal organization of control information as a distinct architectural dimension for locally deployed GUI agents.

Weiming Li, Helen Paik, Yulei Sui · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.