Skip to content
Open access

Study on an Image Candidate Generation Method Based on PCA-Guided Projection

Sep 2026 · Electronics · Vol 15, pp. 3977 · 0 citations · 25 references

TL;DR

The results indicate that PCA-guided projection is an effective lightweight enhancement for systems based on the multi-table p-stable LSH architecture, improving candidate coverage and robustness while remaining compatible with the original online query process.

Abstract

Image retrieval often begins with candidate generation, where a small subset of database images is selected for subsequent reranking or inspection. Standard multi-table p-stable locality-sensitive hashing (LSH) is a classical solution for this stage, but its isotropically sampled projection vectors do not explicitly exploit the principal structure of deep feature distributions. To address this limitation, this paper proposes a PCA-guided candidate generation method using the multi-table bucketization and query architecture of p-stable LSH. Unlike conventional PCA-based feature extraction, PCA is not used here to transform the image features themselves, but to bias the sampling distribution of projection vectors during indexing. Experiments on CALTECH101, CIFAR-10, and Tiny-ImageNet, performed using VGG19, ConvNeXt-Tiny, and ViT-B/16 features, show that the proposed method generally improves on the standard p-stable baseline across different settings. Additional comparisons with HNSW and IVF further position the method relative to modern ANN baselines. The results indicate that PCA-guided projection is an effective lightweight enhancement for systems based on the multi-table p-stable LSH architecture, improving candidate coverage and robustness while remaining compatible with the original online query process.

Read PDF

Similar papers

Open access Sep 2026

Attention-enhanced vision transformer hashing for hybrid image retrieval

Large-scale image retrieval requires compact representations without substantially sacrificing retrieval accuracy. However, Vision Transformer Hashing (VTS) concatenates all output tokens before hash projection, resulting in a high-dimensional hashing head with considerable model and memory overhead. We replace this to...

Uyen Nguyen, Hoai Ba, Quynh Dao Thi Thuy · 0 citations
#edge computing Open access Sep 2026

Fusion of local and global feature representation via optimised transfer learning approach on enhanced content-based image retrieval systems

An Enhanced Content-Based Image Retrieval Using Fusion of Feature Representation with Optimised Image Similarity Measures (CBIRFR-OISM) approach is proposed to effectively enhance the system’s capability to retrieve visually and contextually similar images.

Ravindar Karampuri, Sushama Rani Dutta · 0 citations
Open access Aug 2026

Few-shot image classification algorithm based on deep learning and feature fusion

Few-shot image classification remains difficult because a model must identify novel classes from only one or a few labeled examples while preserving discriminative local information. Metric-learning methods based on Earth Mover’s Distance (EMD) improve local correspondence by representing an image as a set of regional...

Huie Zhang, Mary Jane C. Samontet · 0 citations
Review Aug 2026

Image Augmentation as Test Generation for Deep Learning-Based Image Retrieval Systems

The findings provide practical guidelines for selecting augmentation techniques that maximize test diversity while preserving realistic image characteristics, thereby enabling the construction of comprehensive and effective test suites for image retrieval systems while reducing the cost of manual data labeling through...

Yehan De Silva, Anirudh Sridhar, Armin Lotfy et al. · 0 citations
Open access Sep 2026

Scalable iris image retrieval using attention-based CNNs and locality sensitive hashing (LSH)

Efficient and accurate retrieval of iris images from medium-scale databases plays an important role in national identification, access control, and identity verification. Traditional methods often struggle with scalability, memory demands, and latency as the dataset size grows. We present a scalable retrieval framework...

Fahimeh Afkhamnia, Farsad Zamani Boroujeni, M. Soltanaghaei · 0 citations
Open access Aug 2026

Deep Metric Learning with MobileNetV2 and Triplet Loss for Efficient Image Retrieval

The results show that training the projection head with triplet loss improves embedding quality while reducing the embedding dimensionality from 1280 to 512, which improves retrieval accuracy, but also reduces query latency by more than 59%.

Hersh M. Hama, Kamaran H. Manguri, S. Omer · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.