Benchmark

OmniToM Benchmark

What is OmniToM?

OmniToM evaluates theory of mind in LLMs by requiring explicit belief modeling, extracting belief propositions and labeling them with seven-dimensional schema…

Released
2026-05-25
Evaluates
General AI, Language & Knowledge, cs.AI
Openness
unknown
Importer
Claire Radar
Review status
unreviewed
Reported scores
0

Source provenance

OmniToM paper

No reported scores are on record for this benchmark yet.

Which sources cite OmniToM?

Related General AI, Language & Knowledge, cs.AI benchmarks