MCF-Net: Multi-Modal Cross-Attention Fusion Network with Difference Convolution for Object Detection
Object detection based on visible-light images faces significant challenges under complex illumination and environmental conditions. Introducing infrared or depth images as complementary modalities can effectively enhance detection performance in such scenarios. However, most existing methods only support bi-modal conf...