Benchmark

Local agent benchmark (tools + loop + coding)

What is Local agent benchmark (tools + loop + coding)?

Handful of tasks for manual testing of local LLM tool calling and agent loop behavior. Benchmark Radar tracks Local agent benchmark (tools + loop + coding)…

Released
2026-09-10
Evaluates
Agents, General AI, Language & Knowledge
Openness
unknown
Importer
Claire Radar
Review status
ai-name-audit-deferred
Reported scores
0

Source provenance

Local agent benchmark (tools + loop + coding) code

No reported scores are on record for this benchmark yet.

Which sources cite Local agent benchmark (tools + loop + coding)?

Related Agents, General AI, Language & Knowledge benchmarks