Benchmark
LLM-Uncertainty-Bench
LLM-Uncertainty-Bench is a new benchmarking approach for LLMs that integrates uncertainty quantification. It spans 5 representative natural language…
- Released
- 2024-01-23
- Categories
- 其他, Other, NeurIPS 2024, 大语言模型, LLM, 事实可靠性, Factual Reliability, 不支持, Unsupported
- Openness
- restricted
- Source
- opencompass_hub
- Reported scores
- 0
No reported scores are on record for this benchmark yet.