Benchmark

AICompanionBench

What is AICompanionBench?

AICompanionBench provides a dataset of 2,123 human-AI companion conversations with safety risk annotations to evaluate LLMs-as-judges for detecting unsafe…

Released
2026-06-03
Evaluates
Robotics & Embodied AI, Safety, Safety & Trustworthiness
Openness
unknown
Importer
Claire Radar
Review status
unreviewed
Reported scores
0

Source provenance

AICompanionBench paper and code

No reported scores are on record for this benchmark yet.

Which sources cite AICompanionBench?

Related Robotics & Embodied AI, Safety, Safety & Trustworthiness benchmarks