VoiceMem: Streaming Dual-Brain Memory for Real-Time Interaction
VoiceMem is a streaming dual-brain memory architecture for duplex speech language models, pairing an informational 'left brain' (schema-entity index with cluster emergence) with an emotional 'right brain' (independent and cross-entity persona nodes with short- and long-horizon emotion attribution), fitting within a 134 ms voice-activity-detection silence window.
Why it matters
HF Daily Papers Aug 27 at 157 upvotes; demonstrates sub-VAD-latency streaming memory with verified gains over Mem0 and the newly released ChatMem-Bench, setting a new bar for memory-aware conversational voice agents.
Importance: 3/5
notable release; official+media confirmation; HF Daily 157 upvotes