Benchmark

LUNAR Benchmark

What is LUNAR?

LUNAR is a benchmark for evaluating LLM personalization from longitudinal app interaction logs across domains like clothing, food, housing, and mobility. It…

Released
2026-08-05
Evaluates
General AI, Language & Knowledge, cs.AI
Openness
unknown
Importer
Claire Radar
Review status
unreviewed
Reported scores
0

Source provenance

LUNAR paper

No reported scores are on record for this benchmark yet.

Which sources cite LUNAR?

Related General AI, Language & Knowledge, cs.AI benchmarks