en

Please fill in your name

Mobile phone format error

Please enter the telephone

Please enter your company name

Please enter your company email

Please enter the data requirement

Successful submission! Thank you for your support.

Format error, Please fill in again

Confirm

The data requirement cannot be less than 5 words and cannot be pure numbers

m.nexdata.datatang.com

Speech Recognition Datasets

Instantly enhance AI model performance with high quality off-the-shelf datasets.

Language

All
20
Arabic
4
Burmese
2
Chinese Dialects
13
English
45
French
11
German
9
Hindi
6
Indonesian
8
Italian
8
Japanese
8
Korean
13
Malay
5
Mandarin
11
Others
49
Portuguese
12
Russian
6
Spanish
14
Thai
8
Vietnamese
7

Data Type

All
20
Dialogue
116
Read
111

50.5 Hours Children Speech Dataset (American English) – Scripted Monologue by Microphone

This dataset contains 50.5 hours of American English speech from children, collected from monologue based on given prompts, covering children’s textbooks, story books, oral language, numbers, letters. Transcribed with text content. Our dataset was collected from extensive and diversify speakers(219 native speakers), geographicly speaking, enhancing model performance in real and complex tasks. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.
child english speech dataset kids speech dataset children speech dataset american english speech dataset american children speech dataset

55 Hours Children Speech Dataset (British English) – Scripted Monologue by Microphone

This dataset contains 55 hours of British English speech from children, collected from monologue based on given scripts, covering educational materials for children, story books, informal language, numbers, alphabet. Transcribed with text content and other attributes. Our dataset was collected from extensive and diversify speakers(201 British children recorded in hi-fi microphone), geographicly speaking, enhancing model performance in real and complex tasks. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.
children speech dataset british english speech dataset british children speech dataset kids speech dataset child english speech dataset

96 Hours - Japanese Children Speech Dataset for ASR & TTS

This dataset contains 96 hours of Japanese speech from children, covers self-media, conversation, live, lecture, variety show and other generic domains, mirrors real-world interactions. Transcribed with text content, speaker's ID, gender, age, accent and other attributes. Our dataset was collected from extensive and diversify speakers(12 years old and younger kids), geographicly speaking, enhancing model performance in real and complex tasks. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.
Japanese children speech dataset Japanese audio dataset Japanese speech dataset Japanese ASR training data Japanese child voice corpus Japanese kid speech dataset

Korean Children Speech Dataset – 393 Hours of Scripted Monologues

This 393-hour Korean Children Speech Dataset consists of scripted monologue recordings from young speakers, captured using smartphones. The speech content includes essays, storytelling, and numeric readings. Each audio file is transcribed and annotated with metadata such as speaker ID, gender, and age. Collected from a geographically diverse group of native Korean-speaking children, this dataset is designed to support training of automatic speech recognition (ASR), text-to-speech (TTS), pronunciation evaluation systems, and educational language models. The dataset has been quality-verified by multiple AI enterprises and is fully compliant with GDPR, CCPA, and PIPL privacy regulations.
Korean children speech dataset Korean child voice dataset kids speech recognition dataset Korean Korean ASR training data for children scripted monologue kids Korean smartphone voice dataset Korean Korean kids TTS dataset Korean educational speech corpus

Hindi Children Speech Dataset – 34 Hours (Real-world Conversation & Monologue)

This dataset contains 34 hours of Hindi children’s speech.The recordings cover self-media, conversations, live talk, lectures, variety show and other generic domains, mirrors real-world interactions. Each utterance is transcribed with text content, speaker's ID, gender, age, accent and other attributes. Our dataset was collected from extensive and diversify speakers(12 years old and younger children), geographicly speaking, enhancing model performance in real and complex tasks. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.
Hindi children speech dataset Hindi kids speech dataset Hindi child speech dataset Hindi children voice dataset Hindi speech dataset Hindi children ASR dataset Hindi children TTS dataset

299 Hours – US English Children Speech Dataset for ASR & TTS

This dataset includes 299 hours of US English children’s speech, recorded as scripted monologues, collected from monologue based on given scripts, covering essay stories. The data covers a variety of categories, including children's books and textbooks, and is rich in content that aligns with children's language habits.Transcribed with text content and other attributes. Our dataset is collected from a wide and diverse range of speakers geographically, which supports tasks like speech recognition, TTS, and child voice modeling. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.
children speech dataset kids voice dataset child speech corpus US English children dataset children TTS dataset child speech recognition data American English children voice dataset

97 Hours – German Children Speech Dataset (Conversations & Monologues)

The 97-hour German Children Speech Dataset covers self-media, conversation, live, lecture, variety show and other generic domains, mirrors real-world interactions. Transcribed with text content, speaker's ID, gender, age, accent and other attributes. Our dataset was collected from extensive and diversify speakers(12 years old and younger children), geographicly speaking, enhancing model performance in real and complex tasks. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.
German children speech dataset German kids speech recognition German child speech corpus German ASR dataset children German kids voice dataset German conversational speech children German child dialogue dataset German children NLP dataset German child language dataset multilingual children speech data

189 Hours - Spanish(Latin America) Children Real-world Casual Conversation and Monologue speech dataset

Spanish(Latin America) Children Real-world Casual Conversation and Monologue speech dataset, covers self-media, conversation, live, lecture, variety show and other generic domains, mirrors real-world interactions. Transcribed with text content, speaker's ID, gender, age, accent and other attributes. Our dataset was collected from extensive and diversify speakers(12 years old and younger children), geographicly speaking, enhancing model performance in real and complex tasks. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.
Spanish Casual Conversation Monologue Asr

145 Hours - Spanish(spain) Children Real-world Casual Conversation and Monologue speech dataset

Spanish(spain) Children Real-world Casual Conversation and Monologue speech dataset, covers self-media, conversation, live, lecture, variety show and other generic domains, mirrors real-world interactions. Transcribed with text content, speaker's ID, gender, age, accent and other attributes. Our dataset was collected from extensive and diversify speakers(12 years old and younger children), geographicly speaking, enhancing model performance in real and complex tasks. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.
Spanish Casual Conversation Monologue Asr

loading

Tailor Your Data Now

Why off-the-shelf Datasets

  • Copyright

    Copyright

    Clear Coyright and Ready to Check
  • Security

    Security

    Properly Authorized Secure to Use
  • Professional

    Professional

    Designed and produced by AI data experts
  • Diversity

    Diversity

    Collected from a varity of real scenes
  • Cost Effective

    Cost Effective

    More Cost-Efficient Than Tailored Data
  • Efficiency

    Efficiency

    Ready-To-Go Deliver in Seconds
ac61e154-cdcd-4cec-b816-f0547e0183da