Benchmark

AgentS4D Benchmark

What is AgentS4D?

AgentS4D evaluates runtime safety of LLM-based workspace agents across a four-dimensional framework, with 328 risk-injected cases and seven lifecycle…

Released
2026-07-29
Evaluates
General AI, Safety, Safety & Trustworthiness
Openness
unknown
Importer
Claire Radar
Review status
unreviewed
Reported scores
0

Source provenance

AgentS4D paper

No reported scores are on record for this benchmark yet.

Which sources cite AgentS4D?

Related General AI, Safety, Safety & Trustworthiness benchmarks