Few-shot learning (FSL) provides a promising solution for reducing annotation cost in hyperspectral image change detection (HSI-CD). However, most existing FSL approaches focus on single-modality or homogeneous domains, limiting their ability to handle domain shift. Moreover, many feature extractors rely on fixed kerne...
Hong-Min Gao, Shu-Yu Fei, Shu-Fang Xu et al.· IEEE Transactions on Image P...· 0 citations
Deep unfolding networks (DUNs) offer an iterative paradigm that unrolls optimization procedures into a cascaded network structure for hyperspectral image (HSI) denoising. However, existing DUNs for HSI denoising suffer from two notable limitations: 1) the ill-posed inverse problem of handling severely degraded observat...
Zhen-Zhou Wei, Yu-Bang Zheng, Heng-Chao Li et al.· IEEE Transactions on Geoscie...· 0 citations
Joint classification of multimodal remote sensing data, such as hyperspectral image (HSI) and light detection and ranging (LiDAR), is important for Earth observation. Early deep learning-based methods usually adopt a single network to extract the features of HSI and LiDAR (HSI-LiDAR) data, respectively. However, this k...
Bin Zhao, Wen-Jie Cao, Lin Han et al.· IEEE Transactions on Geoscie...· 0 citations
Tiny object detection (TOD) in remote sensing imagery remains challenging because foreground signals are extremely weak in deep feature hierarchies and are easily overwhelmed by high-response background interference. To mitigate this observed foreground-background signal modulation imbalance (FBSMI) difficulty, we prop...
Tian-Wei Zhang, Longfei Ren, Lian-Ru Gao et al.· IEEE Transactions on Image P...· 0 citations
This work proposes EarthLD, a vision-language-guided diffusion framework for open-world landslide understanding, enabling unified landslide recognition, mapping, and trigger interpretation, and constructs a global-scale open-world landslide benchmark.
Yuan-Chao Su, Lian-Ru Gao, Mengying Jiang et al.· 0 citations
A unified post-training framework that equips pretrained text-to-speech models with natural-language control over segment-level emotion and duration is proposed, highlighting post-training as a practical approach to extending existing speech synthesis models.
Lian-Ru Gao, Yu-Jie Guo, Yong Qin· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.