Benchmark
EnterpriseRAG Benchmark
What is EnterpriseRAG?
The benchmark evaluates LLM instruction adherence and robustness in enterprise retrieval scenarios, using 983 expert-validated samples across six domains…
- Released
- 2026-08-12
- Evaluates
- Factuality, General AI, Information retrieval, Robustness, Safety & Trustworthiness, cs.AI
- Openness
- unknown
- Importer
- Claire Radar
- Review status
- unreviewed
- Reported scores
- 0
Source provenance
- Original evidence https://arxiv.org/abs/2608.11584
EnterpriseRAG paper
No reported scores are on record for this benchmark yet.