Benchmark

FrontierCode

Cognition's production-standard coding evaluation, asking whether models write *good* code rather than merely correct code. Unrelated to Terminal-Bench's…

Released
2026-06-08
Categories
coding_agent
Openness
unknown
Source
model_reports
Reported scores
0

No reported scores are on record for this benchmark yet.