Benchmark

SMH-Bench

What is SMH-Bench?

Benchmark for LLM agents in smart-home environments, with 1,100 tasks across 7 categories and 22 subcategories, built on the executable HomeEnv simulator…

Released
2026-06-01
Evaluates
Agents & Tool Use, General AI, Reasoning
Openness
unknown
Importer
Claire Radar
Review status
unreviewed
Reported scores
0

Source provenance

SMH-Bench paper

No reported scores are on record for this benchmark yet.

Which sources cite SMH-Bench?

Related Agents & Tool Use, General AI, Reasoning benchmarks