Benchmark

TimeConflict Benchmark

What is TimeConflict?

TimeConflict evaluates whether model edits preserve historically correct answers while making updated facts current in the appropriate temporal contexts.

Released
2026-07-13
Evaluates
General AI, Language & Knowledge, cs.LG
Openness
unknown
Importer
Claire Radar
Review status
ai-reviewed
Reported scores
0

Source provenance

TimeConflict paper and code

No reported scores are on record for this benchmark yet.

Which sources cite TimeConflict?

Related General AI, Language & Knowledge, cs.LG benchmarks