VoiceMem: Streaming Dual-Brain Memory for Real-Time Interaction

Research official + media 2 src. ~1 min

VoiceMem is a streaming dual-brain memory architecture for duplex speech language models, pairing an informational 'left brain' (schema-entity index with cluster emergence) with an emotional 'right brain' (independent and cross-entity persona nodes with short- and long-horizon emotion attribution), fitting within a 134 ms voice-activity-detection silence window.

Why it matters

HF Daily Papers Aug 27 at 157 upvotes; demonstrates sub-VAD-latency streaming memory with verified gains over Mem0 and the newly released ChatMem-Bench, setting a new bar for memory-aware conversational voice agents.

Importance: 3/5

notable release; official+media confirmation; HF Daily 157 upvotes

Sources