Benchmark

SEVRA-BENCH

What is SEVRA-BENCH?

SEVRA-BENCH is a benchmark for measuring how often LLM-based code review agents approve adversarial pull requests with social-engineering framings, built from…

Released
2026-06-11
Evaluates
Agents, Cybersecurity, Language & Knowledge, Robotics & Embodied AI
Openness
unknown
Importer
Claire Radar
Review status
unreviewed
Reported scores
0

Source provenance

SEVRA-BENCH paper

No reported scores are on record for this benchmark yet.

Which sources cite SEVRA-BENCH?

Related Agents, Cybersecurity, Language & Knowledge, Robotics & Embodied AI benchmarks