Benchmark
UniHall Benchmark
What is UniHall?
UniHall evaluates multimodal-model hallucinations across object, instruction, and knowledge dimensions, alongside adaptive fuzzing stress tests.
- Released
- 2026-07-15
- Evaluates
- General AI, Multimodal
- Openness
- unknown
- Importer
- Claire Radar
- Review status
- ai-reviewed
- Reported scores
- 0
Source provenance
- Original evidence https://arxiv.org/abs/2608.07525
UniHall paper and code
No reported scores are on record for this benchmark yet.