Benchmark

MMPCBench

What is MMPCBench?

Evaluates multimodal large language models on proactive critique across 3,146 instances with a taxonomy of 4 error types and 12 subtypes, measuring detection…

Released
2026-08-29
Evaluates
General AI, Multimodal
Openness
unknown
Importer
Claire Radar
Review status
ai-reviewed
Reported scores
0

Source provenance

MMPCBench paper and code

No reported scores are on record for this benchmark yet.

Which sources cite MMPCBench?

Related General AI, Multimodal benchmarks