Benchmark

RPCBench

What is RPCBench?

Evaluates LLM-based recommendation assistants on detecting, localizing, and handling faulty premises in natural-language recommendation requests.

Released
2026-09-01
Evaluates
General AI, Language & Knowledge, cs.AI
Openness
unknown
Importer
Claire Radar
Review status
ai-reviewed
Reported scores
0

Source provenance

RPCBench paper and code

No reported scores are on record for this benchmark yet.

Which sources cite RPCBench?

Related General AI, Language & Knowledge, cs.AI benchmarks