Benchmark
SAKE Benchmark
What is SAKE?
SAKE is a standardized multiple-choice evaluation of software-architecture knowledge in language models, covering eight architectural categories and four…
- Released
- 2026-06-28
- Evaluates
- Code & Software, Factuality, General AI, cs.SE
- Openness
- unknown
- Importer
- Claire Radar
- Review status
- ai-reviewed
- Reported scores
- 0
Source provenance
- Original evidence https://arxiv.org/abs/2606.29520
SAKE paper
No reported scores are on record for this benchmark yet.