Benchmark

LatentMD Benchmark

What is LatentMD?

Evaluates Markdown boundary failures, specifically CommonMark-level fence-boundary accuracy, across 4,179 prompts using a scoring CLI that separates content…

Released
2026-09-07
Evaluates
Code & Software, General AI, cs.SE
Openness
unknown
Importer
Claire Radar
Review status
ai-reviewed
Reported scores
0

Source provenance

LatentMD paper

No reported scores are on record for this benchmark yet.

Which sources cite LatentMD?

Related Code & Software, General AI, cs.SE benchmarks