Benchmark
Multi-SWE-Bench
What is Multi-SWE-Bench?
Multilingual SWE-bench variant spanning 7 programming languages; cross-repository context transfer is the measured capability.
- Evaluates
- coding agent
- Openness
- unknown
- Catalog source
- Model reports
- Reported scores
- 2
Multi-SWE-Bench code
Multi-SWE-Bench results and reported scores
Model reports
| Model | Organization | Reported value | Reported | Evidence |
|---|---|---|---|---|
| Seed2.0 Lite | ByteDance | 41.1 | 2026-02-14 | evidence |
| Seed2.0 Pro | ByteDance | 45.2 | 2026-02-14 | evidence |
Scores are partitioned by the source that reported them and are never merged into a single cross-source ranking, because the sources measure different things and say so.
Which sources cite Multi-SWE-Bench?
- Seed2.0 (Pro / Lite / Mini) ByteDance, 2026-02-14