Exploiting Target Knowledge from MLLMs for Robust Few-Shot Segmentation
A novel framework that mines target knowledge using the strong reasoning capacity of Multimodal Large Language Models (MLLMs) and employs it to enhance FSS, which shows promising results and largely surpasses existing methods.