Skip to content
Review

Image Augmentation as Test Generation for Deep Learning-Based Image Retrieval Systems

Aug 2026 · 0 citations · 93 references
Computer Science

TL;DR

The findings provide practical guidelines for selecting augmentation techniques that maximize test diversity while preserving realistic image characteristics, thereby enabling the construction of comprehensive and effective test suites for image retrieval systems while reducing the cost of manual data labeling through the use of metamorphic testing.

Abstract

Ensuring the reliability of deep learning-based image retrieval systems is a software engineering challenge. This paper presents a dual contribution: (1) a literature review of augmentation and generation techniques which resulted in the identification of 50 techniques which we organized into a ten-category taxonomy, and (2) a large-scale empirical study that evaluates these techniques as test generators for embedding-based image retrieval systems. Augmented images are embedded using Amazon Titan and OpenCLIP, and evaluated across four analytical dimensions: (1) embedding-space similarity, (2) embedding uncertainty measured via four estimators, (3) semantic realism scored by LLaVA, and (4) retrieval failure rate. Experiments are performed on three datasets: CIFAR-10, ImageNet-1K, and a dataset from an industrial partner (March Networks). Across all evaluated datasets and embedding models, and under the single severity level tested for each technique, weather simulation and SaSPA are the image augmentation/generation techniques that produce the highest embedding uncertainty and failure rates while maintaining a favorable balance between performance stability, visual realism, and augmentation effectiveness. The results we discuss are configuration-specific and may shift under milder or stronger perturbation settings. In contrast, GAN-based augmentation techniques are among the lowest in realism, indicating the presence of synthetic artifacts and perceptual inconsistencies that reduce their suitability to produce realistic test inputs. Overall, our findings provide practical guidelines for selecting augmentation techniques that maximize test diversity while preserving realistic image characteristics, thereby enabling the construction of comprehensive and effective test suites for image retrieval systems while reducing the cost of manual data labeling through the use of metamorphic testing.

View source

Similar papers

Open access Sep 2026

A Deep Mixed-Image Augmentation Strategy for Few-Shot Image Classification

A deep mixed data augmentation framework that jointly enhances both the support set and the query set in few-shot image classification and is competitive in few-shot classification.

Rui Wang, Xiao-Min Liu · 0 citations
Preprint Aug 2026

Test-Time Augmentation for Tabular-to-Image Classifiers under Distribution Shifts

Results indicate that TTA improves OOD performance, with composite and photometric strategies providing the best trade-off between robustness and variance, in contrast to frequency-domain transformations that alter the encoder's feature-to-intensity mapping consistently degrade performance.

Malena Loza, Felipe Grijalva, Eva Milara et al. · 0 citations
Open access Aug 2026

Deep Metric Learning with MobileNetV2 and Triplet Loss for Efficient Image Retrieval

The results show that training the projection head with triplet loss improves embedding quality while reducing the embedding dimensionality from 1280 to 512, which improves retrieval accuracy, but also reduces query latency by more than 59%.

Hersh M. Hama, Kamaran H. Manguri, S. Omer · 0 citations
Open access Jul 2026

Attention-Based Deep Learning Pipeline for AI-Created Image Recognition

The proposed Attention-Based Deep Learning Pipeline of AI-Created Image Recognition incorporates three integrated branches, including low-level statistical feature extraction, high-level semantic representation learning, and attention-based feature refinement mechanism, which support the robustness and generalization a...

Nadia Ali · 0 citations
Review Open access Jul 2026

Text-to-Image Generation via Deep Learning: A Comprehensive Review of Models, Architectures, and Future Directions

The paradigms of the generative adversarial networks (GANs), variational autoencoders (VAEs), transformer-based designs, and diffusion models, with the last one representing the state of the art in image generation models are discussed.

Abdussalam Elhanashi, Siham Essahraui, Qinghe Zheng et al. · 0 citations
Open access Sep 2026

Study on an Image Candidate Generation Method Based on PCA-Guided Projection

The results indicate that PCA-guided projection is an effective lightweight enhancement for systems based on the multi-table p-stable LSH architecture, improving candidate coverage and robustness while remaining compatible with the original online query process.

Yang-Tay Sun, Tian-Qi Wu, Mao-Lin Li et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.