en

Please fill in your name

Mobile phone format error

Please enter the telephone

Please enter your company name

Please enter your company email

Please enter the data requirement

Successful submission! Thank you for your support.

Format error, Please fill in again

Confirm

The data requirement cannot be less than 5 words and cannot be pure numbers

m.nexdata.datatang.com

Speech Recognition Datasets

Instantly enhance AI model performance with high quality off-the-shelf datasets.

Language

All
231
Arabic
4
Burmese
2
Chinese Dialects
14
English
45
French
11
German
9
Hindi
6
Indonesian
8
Italian
9
Japanese
8
Korean
13
Malay
5
Mandarin
14
Others
49
Portuguese
12
Russian
6
Spanish
14
Thai
8
Vietnamese
7

Data Type

All
231
Dialogue
120
Read
112

200 Hours - Japanese Full-Duplex Multi-Channel Speech Dataset

200 Hours Japanese Full-Duplex Multi-Channel Speech Dataset is collected from dialogues based on given topics. Transcribed with text content, speaker's ID, gender, age and other attributes. Our dataset was collected from extensive and diversify speakers, geographicly speaking, enhancing model performance in real and complex tasks. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.
Japanese speech dataset spontaneous Japanese dialogue multi-stream Japanese audio data conversational Japanese corpus Japanese voice dataset full-duplex speech dataset multi-stream speech dataset multi-channel audio dataset

600 Hours - American English Full-Duplex Multi-Channel Speech Dataset

600 Hours - American English Full-Duplex Multi-Channel Speech Dataset, collected from dialogues based on given topics. Transcribed with text content, speaker's ID, gender, age and other attributes. Our dataset was collected from extensive and diversify speakers, geographicly speaking, enhancing model performance in real and complex tasks. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.
American English speech dataset multi-stream speech dataset full-duplex dialogue dataset spontaneous speech dataset smartphone speech data multi-channel audio dataset speech recognition training data dialogue AI dataset

200 Hours Korean Full-Duplex Multi-Channel Speech Dataset

This 200 Hours Korean Full-Duplex Multi-Channel Speech Dataset features multi-stream audio recorded via smartphones, simulating natural conversations across a range of everyday topics. Each dialogue is annotated with transcripts, speaker ID, gender, and age. Collected from diverse speakers across various regions in Korea, the dataset enhances AI model robustness in real-world applications. Ideal for training automatic speech recognition (ASR), conversational AI, multilingual dialogue systems, and natural speech processing models. All data complies with GDPR, CCPA, and PIPL privacy standards, ensuring safe and ethical AI training.
Korean speech dataset spontaneous dialogue Korean multi-stream audio dataset conversational Korean speech smartphone-recorded audio dual-speaker dataset real-world Korean conversation full-duplex speech dataset

351 People - German Scripted Monologue Speech Dataset (Smart Car & Smart Home)

351 People - German Scripted Monologue Speech Dataset collected from monologue based on given prompts, covering smart car, smart home, voice assistant domains. Transcribed with text content and other attributes. Our dataset was collected from extensive and diversify speakers(351 native speakers), geographicly speaking, enhancing model performance in real and complex tasks. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.
German speech data German voice data German scripted speech data german scripted monologue speech dataset german smartphone speech data german voice assistant speech data

268 Hours - Arabic(Saudi) Full-Duplex Multi-Channel Customer Service Speech Dataset

268 Hours - Arabic(Saudi) Full-Duplex Multi-Channel Customer Service Speech Dataset. Transcribed with text content, speaker's ID, gender, age and other attributes. Our dataset was collected from extensive and diversify speakers(268 native speakers), geographicly speaking, enhancing model performance in real and complex tasks. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.
full-duplex speech dataset multi-channel audio dataset multi-stream speech dataset Saudi Arabic speech dataset Arabic customer service speech dataset full-duplex multi-channel speech dataset call center Arabic speech data

950 Hours - Tagalog Full-Duplex Multi-Channel Speech Dataset

950 Hours - Tagalog Full-Duplex Multi-Channel Speech Dataset, collected from dialogues based on given topics. Transcribed with text content, speaker's ID, gender, age and other attributes. Our dataset was collected from extensive and diversify speakers, geographicly speaking, enhancing model performance in real and complex tasks. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.
multi-stream speech dataset full-duplex dialogue dataset multi-channel audio dataset speech recognition training data multi-channel voice dataset

4600 Hours - Mandarin Full-Duplex Multi-Channel Speech Dataset

4600 Hours Mandarin Full-Duplex Multi-Channel Speech Dataset is collected from dialogues based on given topics. Transcribed with text content, speaker's ID, gender, age and other attributes. Our dataset was collected from extensive and diversify speakers, geographicly speaking, enhancing model performance in real and complex tasks. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.
Mandarin speech dataset multi-stream Mandarin audio data conversational Mandarin corpus Chinese voice dataset full-duplex speech dataset multi-stream speech dataset multi-channel audio dataset

600 Hours - English(Philippine) Full-Duplex Multi-Channel Speech Dataset

600 Hours - English(Philippine) Full-Duplex Multi-Channel Speech Dataset, collected from dialogues based on given topics. Transcribed with text content, speaker's ID, gender, age and other attributes. Our dataset was collected from extensive and diversify speakers, geographicly speaking, enhancing model performance in real and complex tasks. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.
dialogue AI dataset speech recognition training data multi-channel audio dataset smartphone speech data spontaneous speech dataset multi-stream speech dataset Philippine English speech dataset full-duplex speech dataset

791 Hours of Multi-Channel Far-Field Mandarin Conversation Speech Data by Mobile Phone

791 Hours of Multi-Channel Far-Field Mandarin Conversation Speech Data by Mobile Phone, collected from dialogues based on given topics, covering dozens of generic domain. Transcribed with text content, speaker's ID, gender and other attributes. Our dataset was collected from extensive and diversify speakers(1,126 people in total), geographicly speaking, enhancing model performance in real and complex tasks. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.
Mandarin Conversation Far-field voice
. . .

loading

Tailor Your Data Now

Why off-the-shelf Datasets

  • Copyright

    Copyright

    Clear Coyright and Ready to Check
  • Security

    Security

    Properly Authorized Secure to Use
  • Professional

    Professional

    Designed and produced by AI data experts
  • Diversity

    Diversity

    Collected from a varity of real scenes
  • Cost Effective

    Cost Effective

    More Cost-Efficient Than Tailored Data
  • Efficiency

    Efficiency

    Ready-To-Go Deliver in Seconds
0adbc013-f59b-448e-946d-b38732583ebf