Benchmark

SAEScientist-Bench

What is SAEScientist-Bench?

Evaluates AI agents performing sparse autoencoder feature interpretation on Gemma-2-9B-IT, measuring feature rank, activation separation, and steering ability…

Released
2026-09-08
Evaluates
Agents & Tool Use, General AI, cs.AI
Openness
unknown
Importer
Claire Radar
Review status
ai-reviewed
Reported scores
0

Source provenance

SAEScientist-Bench paper and code

No reported scores are on record for this benchmark yet.

Which sources cite SAEScientist-Bench?

Related Agents & Tool Use, General AI, cs.AI benchmarks