Benchmark

AA-Briefcase

AA-Briefcase is an Artificial Analysis evaluation of AI systems on professional knowledge-work tasks, reported as an Elo score.

Modality
text
Categories
productivity, reasoning, agents
Openness
unknown
Source
llm_stats
Reported scores
3

Reported scores

llm_stats

ModelOrganizationReported valueReportedEvidence
Grok 4.6xAI1577.02026-08-12evidence
Inkling-SmallThinking Machines Lab917.02026-07-30evidence
Kimi K3Moonshot AI1548.02026-07-16evidence

Scores are partitioned by the source that reported them and are never merged into a single cross-source ranking, because the sources measure different things and say so.