Benchmark

SWE-Together Benchmark

What is SWE-Together?

SWE-Together evaluates coding agents in multi-turn interactive user sessions reconstructed from real user-agent interactions. It comprises 109 repository-level…

Released
2026-06-29
Evaluates
Agents, Agents & Tool Use, Software & AI Compute
Openness
unknown
Importer
Claire Radar
Review status
unreviewed
Reported scores
0

Source provenance

SWE-Together paper and code

No reported scores are on record for this benchmark yet.

Which sources cite SWE-Together?

Related Agents, Agents & Tool Use, Software & AI Compute benchmarks