Benchmark
FrontierCode
Cognition's production-standard coding evaluation, asking whether models write *good* code rather than merely correct code. Unrelated to Terminal-Bench's…
- Released
- 2026-06-08
- Categories
- coding_agent
- Openness
- unknown
- Source
- model_reports
- Reported scores
- 0
No reported scores are on record for this benchmark yet.