Skip to content

UvA-DARE (Digital Academic Repository) Unsupervised Corpus Poisoning Attacks in Continuous Space for Dense Retrieval

· 1 citation · 49 references

TL;DR

This paper trains a perturbation model with the objective of maintaining the geometric distance between the original and adversarial document embeddings, while also maximizing the token-level dissimilarity be-tween the original and adversarial documents.

View source

Similar papers

#artificial intelligence Preprint Sep 2026

Micro-Collaborative Poisoning: A Distributed Attack on RAG Systems

Retrieval-Augmented Generation (RAG) improves large language models by grounding outputs in external knowledge sources, but this dependency also creates a surface for poisoning attacks. This paper introduces Micro-Collaborative Poisoning, a distributed attack in which a false target claim is divided across multiple loc...

Pedro Pereira, Eva Maia, Isabel Praça · 0 citations
Preprint Aug 2026

DSPrompt: Dynamic Soft Prompt Defense Against M-RAG Corruption

DSPrompt is proposed, a Dynamic Soft Prompt defense framework that directly reshapes the retriever's embedding semantics, without modifying the retrieval pipeline, and is consistently outperforming existing defense baselines at a fraction of their computational cost.

Chang Liu, Y. Lai, Ming-Yue Cui et al. · 0 citations
Aug 2026

Deep Supervised Adversarial Robust Hashing for Retrieval.

Deep Supervised Adversarial Robust Hashing (DSARH), an end-to-end framework that leverages similarity matrices and learnable hash codes to construct gradient-based worst-case perturbations, enabling efficient adversarial training and robust feature learning for retrieval, is proposed.

Xing-Wei Zhang, Gang Zhou, Xiaolong Zheng et al. · 0 citations
Open access Sep 2026

Adversarial Robustness in URL-Based Phishing Detection: Problem-Space Evaluation and Robust Feature Engineering

Machine learning has become a widely adopted approach for URL-based phishing detection, with many studies reporting F1 scores exceeding 0.95 on benchmark datasets. However, recent adversarial machine learning research has questioned the robustness of these models, suggesting that small input perturbations can severely...

Merve Yıldırım · 0 citations
Book Open access Aug 2026

Black-Box Embedding Inversion Attack on Vector Databases

A novel black-box image embedding inversion attack that reconstructs high-fidelity images using only query access to the embedding model or API, and introduces an embedding-guided cross-attention mechanism, where image embeddings serve as conditional signals to steer the generation process.

Lichao Sun, Yun-Cheng Wu, Hai-Chao Sha et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.