SpeakerMem-R1: Speaker-Centered Dual-Track Memory for Multi-Party Dialogue
This research addresses message attribution and relational understanding bottlenecks in multi-party dialogues by proposing SpeakerMem-R1, a dual-track memory system storing speaker-labeled verbatim messages and derived states organized into person-level and group-level views. To reduce attribution and update errors, they train Writer-R1 using SpeakerLevenshtein and speaker-conditioned GRPO. Results show binary accuracies of 47.9%, 69.2%, and 61.9% on GroupMemBench, SocialMemBench, and EverMemBench respectively, alongside 62.33% on the EverMemBench leaderboard and 70.85% on LoCoMo questions.
SpeakerMem-R1 achieves binary accuracies of 47.9%, 69.2%, and 61.9% on GroupMemBench, SocialMemBench, and EverMemBench, respectively.
Achieves 62.33% on the publicly reported EverMemBench leaderboard from EverMind-AI.
It also achieves 70.85% on all 1,986 LoCoMo questions, which we use as a two-person long-term conversation boundary test.