Skip to content
Preprint

CoAnchor: Robust Collaborative Perception under Spatio-Temporal Misalignment via Object-Level Anchors

Aug 2026 · 0 citations · 35 references
Computer Science

TL;DR

This paper proposes CoAnchor, an anchor-centric spatio-temporal alignment framework for asynchronous collaborative perception that builds sparse object-level spatio-temporal anchors as a shared interface for pose correction and tightly connects spatial refinement, temporal propagation, and current-time verification within one unified loop, while keeping the overall correction process lightweight.

Abstract

Collaborative perception extends the sensing range of a single vehicle by fusing observations from nearby agents, which improves the robustness of autonomous driving. In realistic deployments, however, the received collaborator messages are often affected by both communication delay and relative-pose noise, which jointly cause stale observations, spatial misalignment, and unstable feature fusion. Existing methods usually address these issues from either the spatial or temporal side, but handling them jointly in a unified and efficient manner remains challenging. In this paper, we propose CoAnchor, an anchor-centric spatio-temporal alignment framework for asynchronous collaborative perception. Instead of directly reasoning on dense BEV features, CoAnchor builds sparse object-level spatio-temporal anchors as a shared interface for pose correction and tightly connects spatial refinement, temporal propagation, and current-time verification within one unified loop, while keeping the overall correction process lightweight. Extensive experiments on both simulated and real-world datasets illustrate that CoAnchor remains competitive under clean settings and improves the robustness under joint delay and pose perturbations with a favorable practical accuracy-efficiency trade-off.

View source

Similar papers

Oct 2026

Asynchrony-Robust Cooperative Perception and Prediction via Continuous-Time Global State Evolution

Vehicle-to-everything (V2X) collaboration can alleviate the limited perception range and occlusion issues of single-agent autonomous driving. However, most existing cooperative studies still focus on single-frame perception, while the few works on joint cooperative perception and prediction largely rely on fixed-step,...

Han-Xiao Ren, Ke-Qiang Li, Xiang Zhao et al. · 0 citations
Preprint Aug 2026

Towards Collaborative Joint Perception and Prediction: Framework, Baseline Evaluation, and Deployment Perspectives

This work presents a conceptual framework for Collaborative Joint Perception and Prediction (Co-P&P) that improves motion prediction of surrounding road users, thereby enhancing situational awareness in complex and dynamic traffic environments.

Lei Wan, Hannan Ejaz Keen, Alexey V. Vinel · 0 citations
Conference Open access Sep 2026

World4V2X: A Consistency-driven World Model for Robust V2X Cooperative Perception

World4V2X is proposed, the first world model framework tailored for V2X cooperative perception, which introduces a spatial observability modeling module that defines spatial consistency boundaries to distinguish reliable regions from uncertain ones, thereby enabling spatial consistency modeling over multi-agent heterog...

Rui Wang, Shuai Wang, Xiang-Yi Qin et al. · 0 citations
Conference Aug 2026

PrismTrack: Perspective-Aware Multi-Cue Association for Robust Multi-Object Tracking

Multi-Object Tracking (MOT) remains challenging due to object occlusion, complex motions, and detection unreliability in crowded scenarios. We propose an enhanced MOT framework integrating and optimizing state-of-the-art components, specifically Improved Detection Confidence Boost (IDCBoost) and Track-Perspective-Based...

Trung Nghia Huynh, Chi Nhan Huynh, Jia-Ching Wang et al. · 0 citations
Sep 2026

In-Training Masked Reconstruction as Structured Representation Augmentation for Collaboration-Aware V2X Perception

LiDAR-based collaborative perception can mitigate occlusions by exchanging complementary viewpoints via Vehicle-to-Everything (V2X) communication. However, existing methods often depend heavily 3D annotations and adopt a pretraining pipeline that reconstruction serves as initialization. This letter presents a unified m...

Benwu Wang, Xu Li, Xieyuanli Chen et al. · 0 citations
Preprint Aug 2026

CoDS: Robust Collaborative Perception via Expert-driven Detection and BEV Segmentation

Collaborative perception breaks through single-view limitations via multi-agent information exchange. However, multi-source noise such as pose errors and communication delays degrades fusion feature quality, constraining perception performance. Joint training of detection and BEV segmentation provides a natural remedy,...

Jinlong Wang, Yuang Jia, Junhong Lin et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.