Benchmark

agent-model-bench

What is agent-model-bench?

Evaluates language models on tool calling and strict JSON adherence using 32 cases, scoring exact tool and argument matches and JSON parsing with strict and…

Released
2026-08-21
Evaluates
Agents, General AI, Language & Knowledge
Openness
unknown
Importer
Claire Radar
Review status
ai-name-audit-deferred
Reported scores
0

Source provenance

agent-model-bench code

No reported scores are on record for this benchmark yet.

Which sources cite agent-model-bench?

Related Agents, General AI, Language & Knowledge benchmarks