Benchmark

Distract-Bench

What is Distract-Bench?

Distract-Bench evaluates robustness of vision-language models to semantic visual distractions, which are meaningful but task-irrelevant cues that preserve the…

Released
2026-06-08
Evaluates
General AI, Multimodal, Reasoning, Robustness, Safety & Trustworthiness
Openness
unknown
Importer
Claire Radar
Review status
unreviewed
Reported scores
0

Source provenance

Distract-Bench paper and code

No reported scores are on record for this benchmark yet.

Which sources cite Distract-Bench?

Related General AI, Multimodal, Reasoning, Robustness, Safety & Trustworthiness benchmarks