Benchmark

RandomBench

What is RandomBench?

RandomBench evaluates whether multimodal LLMs maintain distributionally neutral behavior when selecting among equivalent options, providing metrics for entropy…

Released
2026-06-04
Evaluates
General AI, Multimodal, Safety & Trustworthiness
Openness
unknown
Importer
Claire Radar
Review status
unreviewed
Reported scores
0

Source provenance

RandomBench paper

No reported scores are on record for this benchmark yet.

Which sources cite RandomBench?

Related General AI, Multimodal, Safety & Trustworthiness benchmarks