VoxMem benchmark tests whether audio models remember who spoke and how
Original title (Chinese)
多会话语音记忆基准VoxMem,专治音频大模型记不住「谁说的、怎么说」
AISummary
VoxMem is a multi-session voice memory benchmark for audio large models, targeting whether models retain who said what and how across conversations. The source text provided is truncated after the first sentence, so no further details on methods, scores, or availability can be reported.
Source: AI Era · aiera.com.cnPublished · added here