Benchmark

IH-Benchmark

What is IH-Benchmark?

IH-Benchmark evaluates instruction-hierarchy robustness in LLMs via conflicting instructions from system, user, and tool outputs, covering 44 constraint…

Released
2026-07-28
Evaluates
General AI, Robustness, Safety & Trustworthiness, cs.CR
Openness
unknown
Importer
Claire Radar
Review status
unreviewed
Reported scores
0

Source provenance

IH-Benchmark paper

No reported scores are on record for this benchmark yet.

Which sources cite IH-Benchmark?

Related General AI, Robustness, Safety & Trustworthiness, cs.CR benchmarks