Benchmark

TREAT Benchmark

What is TREAT?

TREAT evaluates LLMs' ability to recognize theorem identities from equivalence-preserving formula transformations, with 737 identities and 29,480 transformed…

Released
2026-07-29
Evaluates
Factuality, Language & Knowledge, Mathematics & Formal Science, cs.AI
Openness
unknown
Importer
Claire Radar
Review status
unreviewed
Reported scores
0

Source provenance

TREAT paper

No reported scores are on record for this benchmark yet.

Which sources cite TREAT?

Related Factuality, Language & Knowledge, Mathematics & Formal Science, cs.AI benchmarks