Benchmark

Avalon-ToM-Bench

What is Avalon-ToM-Bench?

Avalon-ToM-Bench evaluates fine-grained theory of mind in LLMs using a 2x2 taxonomy of epistemic/motivational reasoning crossed with inference/action, via…

Released
2026-08-10
Evaluates
General AI, Language & Knowledge, cs.AI
Openness
unknown
Importer
Claire Radar
Review status
unreviewed
Reported scores
0

Source provenance

Avalon-ToM-Bench paper

No reported scores are on record for this benchmark yet.

Which sources cite Avalon-ToM-Bench?

Related General AI, Language & Knowledge, cs.AI benchmarks