Benchmark

Do LLMs Trust the Accuser or the Accusation? Measuring Belief Shifts in Werewolf Benchmark

What is Do LLMs Trust the Accuser or the Accusation? Measuring Belief Shifts in Werewolf?

A proposed belief-shift evaluation in Werewolf for analyzing communication skills through belief updating. Benchmark Radar tracks Do LLMs Trust the Accuser or…

Released
2026-09-11
Evaluates
General AI, Language & Knowledge, cs.AI
Openness
unknown
Importer
Claire Radar
Review status
ai-name-audit-deferred
Reported scores
0

Source provenance

Do LLMs Trust the Accuser or the Accusation? Measuring Belief Shifts in Werewolf paper

No reported scores are on record for this benchmark yet.

Which sources cite Do LLMs Trust the Accuser or the Accusation? Measuring Belief Shifts in Werewolf?

Related General AI, Language & Knowledge, cs.AI benchmarks