PROBE: VLM-Guided Discrete Structural Reconfiguration for Customized and Efficient Video Retrieval
Existing self-supervised video hashing methods achieve high efficiency by encoding videos into compact binary representations, but they typically rely on a fixed global similarity geometry that enforces a single notion of similarity across all queries. In many real-world retrieval scenarios, however, the same videos ma...