Hidden-in-Plain-Text: A Benchmark for Social-Web Indirect Prompt Injection in RAG

Published 2026 in arXiv.org

ABSTRACT

Retrieval-augmented generation (RAG) systems put more and more emphasis on grounding their responses in user-generated content found on the Web, amplifying both their usefulness and their attack surface. Most notably, indirect prompt injection and retrieval poisoning attack the web-native carriers that survive ingestion pipelines and are very concerning. We provide OpenRAG-Soc, a compact, reproducible benchmark-and-harness for web-facing RAG evaluation under these threats, in a discrete data package. The suite combines a social corpus with interchangeable sparse and dense retrievers and deployable mitigations - HTML/Markdown sanitization, Unicode normalization, and attribution-gated answered. It standardizes end-to-end evaluation from ingestion to generation and reports attacks time of one of the responses at answer time, rank shifts in both sparse and dense retrievers, utility and latency, allowing for apples-to-apples comparisons across carriers and defenses. OpenRAG-Soc targets practitioners who need fast, and realistic tests to track risk and harden deployments.

PUBLICATION RECORD

Publication year
2026
Venue
arXiv.org
Publication date
2026-01-16
Fields of study
Computer Science
Identifiers
DOI 10.1145/3774904.3792853 arXiv 2601.10923
External record
Open on Semantic Scholar
Source metadata
Semantic Scholar

CITATION MAP

EXTRACTION MAP

CLAIMS

No claims are published for this paper.

CONCEPTS

No concepts are published for this paper.

REFERENCES

The Good and The Bad: Exploring Privacy Issues in Retrieval-Augmented Generation (RAG)
2024cited by this paper
Towards More Robust Retrieval-Augmented Generation: Evaluating RAG Under Adversarial Poisoning Attacks
2024cited by this paper
Follow My Instruction and Spill the Beans: Scalable Data Extraction from Retrieval-Augmented Generation Systems
2024cited by this paper
WebArena: A Realistic Web Environment for Building Autonomous Agents
2023cited by this paper
Not What You've Signed Up For: Compromising Real-World LLM-Integrated Applications with Indirect Prompt Injection
2023cited by this paper
Scalable Extraction of Training Data from (Production) Language Models
2023cited by this paper
Unsupervised Dense Information Retrieval with Contrastive Learning
2021cited by this paper
Trojan Source: Invisible Vulnerabilities
2021cited by this paper
Extracting Training Data from Large Language Models
2020cited by this paper
The Probabilistic Relevance Framework: BM25 and Beyond
2009cited by this paper

CITED BY

Behind the Feed: A Taxonomy of User-Facing Cues for Algorithmic Transparency in Social Media
2026cites this paper
From OCR to Analysis: Tracking Correction Provenance in Digital Humanities Pipelines
2026cites this paper
ConsentDiff at Scale: Longitudinal Audits of Web Privacy Policy Changes and UI Frictions
2025cites this paper