Benchmark

GRACE Benchmark

What is GRACE?

GRACE evaluates step-level faithfulness of chain-of-thought reasoning in context-grounded tasks, with human annotations and a taxonomy of error categories.

Released
2026-06-15
Evaluates
General AI, Language & Knowledge, Reasoning
Openness
unknown
Importer
Claire Radar
Review status
unreviewed
Reported scores
0

Source provenance

GRACE paper

No reported scores are on record for this benchmark yet.

Which sources cite GRACE?

Related General AI, Language & Knowledge, Reasoning benchmarks