InjecAgent

Benchmark of indirect prompt injection cases for tool-integrated agents.

Records citing InjecAgent

Mar 5, 2024
InjecAgent benchmarks indirect prompt injection against tool-integrated LLM agents
DefenseBenchmarkUniversity of Illinois Urbana-Champaign

Zhan, Liang, Ying and Kang release InjecAgent, a benchmark of 1,054 test cases spanning 17 user tools and 62 attacker tools, covering direct harm to users and exfiltration of private data. They evaluate 30 LLM agents and find a ReAct-prompted GPT-4 agent vulnerable in about a quarter of cases.