Benchmark

SWE-NFI Benchmark

What is SWE-NFI?

A benchmark of 188 tasks for evaluating coding agents on non-functional improvements in Python projects, with 92 executable rules combining functional…

Released
2026-07-29
Evaluates
Code & Software, General AI, cs.SE
Openness
unknown
Importer
Claire Radar
Review status
unreviewed
Reported scores
0

Source provenance

SWE-NFI paper

No reported scores are on record for this benchmark yet.

Which sources cite SWE-NFI?

Related Code & Software, General AI, cs.SE benchmarks