Benchmark
AuthorityBench
What is AuthorityBench?
AuthorityBench evaluates how citation-based authority signals affect epistemic behavior in LLMs using 220,564 prompts across four domains, with a 2x2 factorial…
- Released
- 2026-06-11
- Evaluates
- General AI, Safety & Trustworthiness, cs.LG
- Openness
- unknown
- Importer
- Claire Radar
- Review status
- unreviewed
- Reported scores
- 0
Source provenance
- Original evidence https://arxiv.org/abs/2606.13104
AuthorityBench paper and code
No reported scores are on record for this benchmark yet.