Benchmark

BELLS-O Benchmark

What is BELLS-O?

An operational benchmark comparing moderation filters, jailbreak detectors, and general-purpose LLM supervisors across effectiveness, false positives, latency…

Released
2026-06-12
Evaluates
General AI, Language & Knowledge, cs.CR
Openness
unknown
Importer
Claire Radar
Review status
ai-reviewed
Reported scores
0

Source provenance

BELLS-O paper

No reported scores are on record for this benchmark yet.

Which sources cite BELLS-O?

Related General AI, Language & Knowledge, cs.CR benchmarks