Benchmark

RTCbench

What is RTCbench?

Evaluates LLM-generated controllers on simulated closed-loop control tasks using metrics like CVaR@10%, safety gating, and replayable traces.

Released
2026-08-20
Evaluates
General AI, Language & Knowledge, cs.AI
Openness
unknown
Importer
Claire Radar
Review status
ai-name-audit-deferred
Reported scores
0

Source provenance

RTCbench code

No reported scores are on record for this benchmark yet.

Which sources cite RTCbench?

Related General AI, Language & Knowledge, cs.AI benchmarks