Benchmark

HLL Benchmark

What is HLL?

HLL is a benchmark that evaluates multimodal agents on interactive CAPTCHA verification in a closed-loop GUI environment, covering diverse task types such as…

Released
2026-06-01
Evaluates
Agents & Tool Use, General AI, cs.AI
Openness
unknown
Importer
Claire Radar
Review status
unreviewed
Reported scores
0

Source provenance

HLL paper and code

No reported scores are on record for this benchmark yet.

Which sources cite HLL?

Related Agents & Tool Use, General AI, cs.AI benchmarks