Benchmark

AgentRedBench

What is AgentRedBench?

AgentRedBench evaluates LLM agents against indirect prompt injection and underspecified-authorization attacks across 24 enterprise SaaS integrations. It…

Released
2026-06-01
Evaluates
General AI, Language & Knowledge, cs.CR
Openness
unknown
Importer
Claire Radar
Review status
unreviewed
Reported scores
0

Source provenance

AgentRedBench paper

No reported scores are on record for this benchmark yet.

Which sources cite AgentRedBench?

Related General AI, Language & Knowledge, cs.CR benchmarks