Skip to content
AI Atlas
PaperActive

PRAGMA: Evaluating Personalized Guidance with Memory Alignment in Lifelong Conversations

arxiv.org/abs/2609.09664

quality89

Updated 5 h ago · first seen 11 Sept 2026

paper_01M294GK7F4ZKYWNH9TTFMRK20

Published
11 Sept 2026
T1 · 5 h ago
arXiv
2609.09664
T1 · 5 h ago
Category
cs.AI
T1 · 5 h ago

Abstract

Large language models (LLMs) are increasingly deployed as personalized assistants that interact with users over extended periods of time. As conversations grow longer, relying on full interaction histories becomes increasingly inefficient and unreliable: long contexts introduce substantial computational overhead, making it difficult for models to consistently identify and utilize the most relevant information for the current request. These challenges have motivated memory systems that structure and retrieve user-specific information. In realistic interactions, users often seek practical guidance such as recommendations, planning, and decision support. Unlike factual recall tasks, personalized guidance requires models to integrate information across multiple past conversations and reason about changing user preferences and experiences. However, existing conversational memory evaluations mainly focus on retrieval and factual recall. To study this challenge, we introduce PRAGMA, a benchmark for evaluating personalized guidance in long-term conversations. PRGAMA contains curated longitudinal conversation histories, evidence annotations, and guidance scenarios grounded in evolving user contexts and incorrect user assumptions. Experiments across retrieval systems, memory systems, and long-context models reveal that current systems struggle both to recover the appropriate conversational evidence and to effectively use it for personalized guidance. Our results highlight the need for memory architectures that support robust conversational retrieval and memory-grounded reasoning beyond evidence recall.

Authors 5

Hyojeong Yu, Hyukhun Koh, Minsung Kim, Yunah Jang, Kyomin Jung

Specification

Official page

Source:arXiv (Atom API + RSS)T1observed 5 h agohigh

Arxiv announce type
new

Source:arXiv (Atom API + RSS)T1observed 5 h agohigh

arXiv id
2609.09664

Source:arXiv (Atom API + RSS)T1observed 5 h agohigh

Categories
cs.AI

Source:arXiv (Atom API + RSS)T1observed 5 h agohigh

PDF

Source:arXiv (Atom API + RSS)T1observed 5 h agohigh

Primary category
cs.AI

Source:arXiv (Atom API + RSS)T1observed 5 h agohigh

Published
11 Sept 2026

Source:arXiv (Atom API + RSS)T1observed 5 h agohigh

Each value shows its source, tier and observation time. Conflicting claims are kept side by side and flagged — never averaged. How AI Atlas records facts →

Provenance

Attributed facts

9

Source tiers

T19

Freshest observation

5 h ago

Conflicts

None