Benchmark

TAR-Bench

What is TAR-Bench?

Evaluates video-language models on ten traffic anomaly reasoning tasks using 960 human-curated test annotations over 80 held-out clips.

Released
2026-08-10
Evaluates
General AI, Multimodal, Reasoning
Openness
unknown
Importer
Claire Radar
Review status
ai-reviewed
Reported scores
0

Source provenance

TAR-Bench paper and dataset

No reported scores are on record for this benchmark yet.

Which sources cite TAR-Bench?

Related General AI, Multimodal, Reasoning benchmarks