Cross-modal learning for SAR target recognition using optical vision foundation models
This work proposes a cross-modal EO to SAR prototype alignment framework in which a frozen EO encoder, based on a DINOv3 vision foundation model, is used to construct class level optical prototypes without requiring strict EO/SAR pairs.
Lucas Hirsch, James R. Hopgood, Javid Khan et al.
· Artificial Intelligence for... · 0 citations