multilingual-speech-commands-15lang artur-muratov

Multilingual Speech Commands Dataset (15 Languages, Augmented) This dataset contains augmented speech command samples in 15 languages, derived from multiple public datasets. Only commands that overlap with the Google Speech Commands (GSC) vocabulary are included, making the dataset suitable for multilingual keyword spotting tasks aligned with GSC-style classification. Audio samples have been augmented using standard audio techniques to improve model robustness (e.g., time-shifting… See the full description on the dataset page: https://huggingface.co/datasets/artur-muratov/multilingual-speech-commands-15lang.

类型
dataset
许可
cc-by-4.0
语言
en
下载量
584,524
点赞
16
访问
public
文件
0

标签

  • 音频数据集
  • 语音识别
  • 音频分类
  • 多语言
  • 语音
  • 关键词识别

摘要

该数据集包含15种语言的增强语音指令样本,仅保留与Google语音指令(GSC)词汇重叠的指令,适用于多语言关键词识别(KWS)任务。音频样本通过时间偏移、噪声注入、音调变化等标准技术进行增强以提升模型鲁棒性。按指令标签分文件夹组织,并附训练/验证/测试列表及标签映射等元数据,适合用于低资源语言和智能系统语音控制场景。

README

--- license: cc-by-4.0 language: - en - ru - kk - tt - ar - tr - fr - de - es - it - ca - fa - pl - nl - rw pretty_name: Multilingual Speech Commands Dataset (15 Languages, Augmented) tags: - speech - audio - keyword-spotting - speech-commands - multilingual - low-resource - dataset - augmentation …

查看完整页面 · 查看原文