Benchmark

VIABLE Benchmark

What is VIABLE?

A benchmark of over 300K judgment samples that evaluates VLM judges for visually impaired assistance on effectiveness, impartiality, and stability.

Released
2026-05-29
Evaluates
General AI, Language & Knowledge, cs.CL
Openness
unknown
Importer
Claire Radar
Review status
ai-reviewed
Reported scores
0

Source provenance

VIABLE paper and code

No reported scores are on record for this benchmark yet.

Which sources cite VIABLE?

Related General AI, Language & Knowledge, cs.CL benchmarks