arXiv NLP
techCenter
When Users Don't Ask: Benchmarking Context-Driven Memory Retrieval in Conversational Agentstranslating…
arXiv:2609.03467v1 Announce Type: new
Abstract: Large language models (LLMs) are increas- ingly deployed as long-horizon conversational agents, motivating growing interest in mem- ory systems. However, existing benchmarks primarily evaluate memory through QA-style probing…
Keywords#Conversational Agents#Memory#Ask#Benchmarking#Benchmarking Context-Driven
Comments
Sign in to join the discussion.
Loading comments…