Co-Attention Guided Multimodal Fusion for Robust 6-D Pose Estimation in Occluded Scenarios
Achieving accurate and efficient object pose estimation is a key goal in computer vision. Most existing methods rely on controlled environments, limiting their effectiveness in complex, dynamic, and unstructured real-world scenarios, especially for novel objects, severe occlusion, or sensor noise. Recent studies show t...