Benchmark
TAR-Bench
What is TAR-Bench?
Evaluates video-language models on ten traffic anomaly reasoning tasks using 960 human-curated test annotations over 80 held-out clips.
- Released
- 2026-08-10
- Evaluates
- General AI, Multimodal, Reasoning
- Openness
- unknown
- Importer
- Claire Radar
- Review status
- ai-reviewed
- Reported scores
- 0
Source provenance
- Original evidence https://arxiv.org/abs/2608.10317
TAR-Bench paper and dataset
No reported scores are on record for this benchmark yet.