Benchmark

ChildEval

What is ChildEval?

ChildEval is a benchmark for evaluating LLMs' ability to infer and follow child-centered preferences in long-context conversations. It contains 29K synthesized…

Released
2026-05-27
Evaluates
General AI, Language & Knowledge, Long Context
Openness
unknown
Importer
Claire Radar
Review status
unreviewed
Reported scores
0

Source provenance

ChildEval paper and code

No reported scores are on record for this benchmark yet.

Which sources cite ChildEval?

Related General AI, Language & Knowledge, Long Context benchmarks