Common Voice

Introduced by Ardila et al. in Common Voice: A Massively-Multilingual Speech Corpus

Common Voice is an audio dataset that consists of a unique MP3 and corresponding text file. There are 9,283 recorded hours in the dataset. The dataset also includes demographic metadata like age, sex, and accent. The dataset consists of 7,335 validated hours in 60 languages.

Papers


Paper Code Results Date Stars

Dataset Loaders