Google Research·· 2025-12-04AI 评分42
Google 发布 MSEB 基准测试:评估多模态听觉智能的八大核心能力
From Waveforms to Wisdom: The New Benchmark for Auditory Intelligence
AI 导读
Google Research 发布 Massive Sound Embedding Benchmark (MSEB),旨在标准化评估多模态模型的听觉智能。该基准涵盖语音检索、推理、分类等八大核心能力,并包含包含 177,352 条语音查询的 Simple Voice Questions (SVQ) 数据集。实验显示当前声音表示远非通用,在各项任务中均存在显著的性能提升空间。
来源:Google Research · research.google