This dataset was recorded in a quiet office/home setting, with the participation of 104 speakers, comprising 51 males and 53 females. All speakers who took part in the recording were carefully selected by professionals to ensure standard pronunciation and articulate speech. The recorded texts encompass topics such as family, travel, and food.
This corpus contains 5,763 speakers with a balanced gender ratio. The speakers are from Spain, Mexico, America,Argentina, and Colombia. The age range is from 16 to 80 years old.
This corpus covers 12 languages of India with 13,150 speakers.The languages including
Assamese,English,Gujarati,Hindi,Kashmiri,Malayalam,Marathi,Odia,Punjabi,Tamil,Telugu,and Urdu
This corpus comprises recordings from 35,628 speakers with each speaker contributing between 10 to 60 minutes of speech. The gender distribution is approximately equal. The age range of the speakers spans from 7 to 80 years old. It includes a diverse array of accents, representing 64 countries including China, the United States, the United Kingdom, Canada, Australia, Japan, South Korea, and many others.
Morocco Arabic Speech Recognition Corpus ( Phone )
This dataset covers free dialogue content, the topics include news, text messages, car control, music, general, maps, daily oral language, family, health, travel, work, socializing, celebrities, weather, and other common topics in life.