Benchmark

AuthorityBench

What is AuthorityBench?

AuthorityBench evaluates how citation-based authority signals affect epistemic behavior in LLMs using 220,564 prompts across four domains, with a 2x2 factorial…

Released
2026-06-11
Evaluates
General AI, Safety & Trustworthiness, cs.LG
Openness
unknown
Importer
Claire Radar
Review status
unreviewed
Reported scores
0

Source provenance

AuthorityBench paper and code

No reported scores are on record for this benchmark yet.

Which sources cite AuthorityBench?

Related General AI, Safety & Trustworthiness, cs.LG benchmarks