Benchmark

AdversaBench

What is AdversaBench?

AdversaBench is an automated LLM red-teaming evaluation that mutates seed prompts and confirms resulting failures with multiple judges and a tiebreaker.

Released
2026-06-23
Evaluates
Cybersecurity, Language & Knowledge, cs.AI
Openness
unknown
Importer
Claire Radar
Review status
ai-reviewed
Reported scores
0

Source provenance

AdversaBench paper and code

No reported scores are on record for this benchmark yet.

Which sources cite AdversaBench?

Related Cybersecurity, Language & Knowledge, cs.AI benchmarks