Benchmark

SWE-Gate Benchmark

What is SWE-Gate?

Evaluates repository-level repair agents on both functional correctness and adherence to review-derived constraints, using 303 instances with separate…

Released
2026-09-03
Evaluates
Code, Code & Software, General AI, Software & AI Compute
Openness
unknown
Importer
Claire Radar
Review status
ai-reviewed
Reported scores
0

Source provenance

SWE-Gate paper and code

No reported scores are on record for this benchmark yet.

Which sources cite SWE-Gate?

Related Code, Code & Software, General AI, Software & AI Compute benchmarks