Benchmark

HEART-Bench

What is HEART-Bench?

HEART-Bench evaluates whether LLM agents can simulate coherent, human-like psychology. It provides 11 fictional characters with raw episodic memories, 64…

Released
2026-05-28
Evaluates
General AI, Language & Knowledge, cs.CL
Openness
unknown
Importer
Claire Radar
Review status
unreviewed
Reported scores
0

Source provenance

HEART-Bench paper and code

No reported scores are on record for this benchmark yet.

Which sources cite HEART-Bench?

Related General AI, Language & Knowledge, cs.CL benchmarks