D-VLC: Decentralized Vision-Language Collaboration for Heterogeneous Embodied Multi-Robot Systems in Unknown Environments
A framework that combines decentralized asynchronous reasoning, lightweight information sharing, capability aware collaboration, and a unified action interface is proposed, enabling general purpose VLMs to generate robot specific actions executed by learning free experts without task or robot specific training.