Benchmark

CombEval

What is CombEval?

CombEval evaluates combinatorial counting abilities of large language models using problems generated from typed Cofola specifications, with solver-verified…

Released
2026-06-18
Evaluates
General AI, Language & Knowledge, cs.AI
Openness
unknown
Importer
Claire Radar
Review status
unreviewed
Reported scores
0

Source provenance

CombEval paper and code

No reported scores are on record for this benchmark yet.

Which sources cite CombEval?

Related General AI, Language & Knowledge, cs.AI benchmarks