Benchmark
Distract-Bench
What is Distract-Bench?
Distract-Bench evaluates robustness of vision-language models to semantic visual distractions, which are meaningful but task-irrelevant cues that preserve the…
- Released
- 2026-06-08
- Evaluates
- General AI, Multimodal, Reasoning, Robustness, Safety & Trustworthiness
- Openness
- unknown
- Importer
- Claire Radar
- Review status
- unreviewed
- Reported scores
- 0
Source provenance
- Original evidence https://arxiv.org/abs/2606.08894
Distract-Bench paper and code
No reported scores are on record for this benchmark yet.