Benchmark

Blind-Spots-Bench

What is Blind-Spots-Bench?

Evaluates reasoning blind spots in language, vision-language, and image-generation models across 235 samples with structured reference solutions and taxonomy.

Released
2026-07-09
Evaluates
General AI, Multimodal
Openness
unknown
Importer
Claire Radar
Review status
unreviewed
Reported scores
0

Source provenance

Blind-Spots-Bench paper and code

No reported scores are on record for this benchmark yet.

Which sources cite Blind-Spots-Bench?

Related General AI, Multimodal benchmarks