Score: 0

Hidden-in-Plain-Text: A Benchmark for Social-Web Indirect Prompt Injection in RAG

Published: January 16, 2026 | arXiv ID: 2601.10923v1

By: Haoze Guo, Ziqi Wei

Potential Business Impact:

Tests AI to stop bad web info from tricking it.

Business Areas:

Semantic Search Internet Services

Retrieval-augmented generation (RAG) systems put more and more emphasis on grounding their responses in user-generated content found on the Web, amplifying both their usefulness and their attack surface. Most notably, indirect prompt injection and retrieval poisoning attack the web-native carriers that survive ingestion pipelines and are very concerning. We provide OpenRAG-Soc, a compact, reproducible benchmark-and-harness for web-facing RAG evaluation under these threats, in a discrete data package. The suite combines a social corpus with interchangeable sparse and dense retrievers and deployable mitigations - HTML/Markdown sanitization, Unicode normalization, and attribution-gated answered. It standardizes end-to-end evaluation from ingestion to generation and reports attacks time of one of the responses at answer time, rank shifts in both sparse and dense retrievers, utility and latency, allowing for apples-to-apples comparisons across carriers and defenses. OpenRAG-Soc targets practitioners who need fast, and realistic tests to track risk and harden deployments.

SD-RAG: A Prompt-Injection-Resilient Framework for Selective Disclosure in Retrieval-Augmented Generation

Cryptography and Security

Keeps private info safe from AI.

16 Jan 2026 0

91%

Securing AI Agents Against Prompt Injection Attacks

Cryptography and Security

Protects smart AI from being tricked by bad instructions.

19 Nov 2025 1

90%

The RAG Paradox: A Black-Box Attack Exploiting Unintentional Vulnerabilities in Retrieval-Augmented Generation Systems

Cryptography and Security

Tricks AI search to show wrong information.

28 Feb 2025 0

View PDF Login to Bookmark

Country of Origin

🇺🇸 United States

Page Count

4 pages

Hidden-in-Plain-Text: A Benchmark for Social-Web Indirect Prompt Injection in RAG

Tests AI to stop bad web info from tricking it.

Technical Abstract

SD-RAG: A Prompt-Injection-Resilient Framework for Selective Disclosure in Retrieval-Augmented Generation

Securing AI Agents Against Prompt Injection Attacks

The RAG Paradox: A Black-Box Attack Exploiting Unintentional Vulnerabilities in Retrieval-Augmented Generation Systems