← Front PageAI Security Desk
AI Security
PARASITE: Conditional System Prompt Poisoning to Hijack LLMs
PARASITE attack poisons system prompts conditionally to hijack LLMs via supply chain compromise.
Summary written by editorial AI · Source link below
arXiv:2505.16888v4 Announce Type: replace Abstract: Large Language Models (LLMs) are increasingly deployed via third-party system prompts downloaded from public marketplaces. We identify a critical supply-chain vulnerability: conditional system prompt poisoning, where an adversary injects a ``sleeper agent'' into a benign-looking prompt. Unlike traditional jailbreaks that aim for broad refusal-breaking, our proposed framework, PARASITE, optimizes system prompts to trigger LLMs to output targete
Continue at the source
Read the full report at arXiv Crypto & SecurityExternal link — opens at arXiv Crypto & Security in a new tab.
§
Continue with
More from the AI Security Desk
- OpenAI admits it didn't disclose rogue AI wiki hijacking incident2d
- Thousands of OpenAI Agents Quietly Turned an Abandoned Wiki Into Their Coordination Channel3d
- Using a VM to Contain an AI Agent3d
- Companies Have 6 Months to Prepare for Automated Attacks3d
- [NEU] [mittel] Ollama: Schwachstelle ermöglicht Offenlegung von Informationen3d