Benchmark

LLM-Uncertainty-Bench

LLM-Uncertainty-Bench is a new benchmarking approach for LLMs that integrates uncertainty quantification. It spans 5 representative natural language…

Released
2024-01-23
Categories
其他, Other, NeurIPS 2024, 大语言模型, LLM, 事实可靠性, Factual Reliability, 不支持, Unsupported
Openness
restricted
Source
opencompass_hub
Reported scores
0

No reported scores are on record for this benchmark yet.