Benchmark

NovGauge Benchmark

What is NovGauge?

Human-anchored assessment of LLM capabilities in diagnosing paper novelty. Benchmark Radar tracks NovGauge, cited by 1 source documents, with links to its…

Released
2026-09-10
Evaluates
General AI, Language & Knowledge, cs.AI
Openness
unknown
Importer
Claire Radar
Review status
ai-name-audit-deferred
Reported scores
0

Source provenance

NovGauge paper

No reported scores are on record for this benchmark yet.

Which sources cite NovGauge?

Related General AI, Language & Knowledge, cs.AI benchmarks