multilingual-speech-commands-15lang artur-muratov

Multilingual Speech Commands Dataset (15 Languages, Augmented) This dataset contains augmented speech command samples in 15 languages, derived from multiple public datasets. Only commands that overlap with the Google Speech Commands (GSC) vocabulary are included, making the dataset suitable for multilingual keyword spotting tasks aligned with GSC-style classification. Audio samples have been augmented using standard audio techniques to improve model robustness (e.g., time-shifting… See the full description on the dataset page: https://huggingface.co/datasets/artur-muratov/multilingual-speech-commands-15lang.

유형
dataset
라이선스
cc-by-4.0
언어
en
다운로드
584,524
좋아요
16
접근
public
파일
0

README

--- license: cc-by-4.0 language: - en - ru - kk - tt - ar - tr - fr - de - es - it - ca - fa - pl - nl - rw pretty_name: Multilingual Speech Commands Dataset (15 Languages, Augmented) tags: - speech - audio - keyword-spotting - speech-commands - multilingual - low-resource - dataset - augmentation …

查看完整页面 · 查看原文