en

Please fill in your name

Mobile phone format error

Please enter the telephone

Please enter your company name

Please enter your company email

Please enter the data requirement

Successful submission! Thank you for your support.

Format error, Please fill in again

Confirm

The data requirement cannot be less than 5 words and cannot be pure numbers

m.nexdata.datatang.com

Speech Recognition Datasets

Instantly enhance AI model performance with high quality off-the-shelf datasets.

Language

All
225
Arabic
4
Burmese
2
Chinese Dialects
13
English
46
French
11
German
9
Hindi
6
Indonesian
8
Italian
8
Japanese
8
Korean
13
Malay
5
Mandarin
11
Others
48
Portuguese
12
Russian
6
Spanish
14
Thai
8
Vietnamese
7

Data Type

All
225
Dialogue
116
Read
110

600 Hours - American English Full-Duplex Multi-Channel Speech Dataset

600 Hours - American English Full-Duplex Multi-Channel Speech Dataset, collected from dialogues based on given topics. Transcribed with text content, speaker's ID, gender, age and other attributes. Our dataset was collected from extensive and diversify speakers, geographicly speaking, enhancing model performance in real and complex tasks. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.
American English speech dataset multi-stream speech dataset full-duplex dialogue dataset spontaneous speech dataset smartphone speech data multi-channel audio dataset speech recognition training data dialogue AI dataset

200 Hours Korean Full-Duplex Multi-Channel Speech Dataset

This 200 Hours Korean Full-Duplex Multi-Channel Speech Dataset features multi-stream audio recorded via smartphones, simulating natural conversations across a range of everyday topics. Each dialogue is annotated with transcripts, speaker ID, gender, and age. Collected from diverse speakers across various regions in Korea, the dataset enhances AI model robustness in real-world applications. Ideal for training automatic speech recognition (ASR), conversational AI, multilingual dialogue systems, and natural speech processing models. All data complies with GDPR, CCPA, and PIPL privacy standards, ensuring safe and ethical AI training.
Korean speech dataset spontaneous dialogue Korean multi-stream audio dataset conversational Korean speech smartphone-recorded audio dual-speaker dataset real-world Korean conversation full-duplex speech dataset

351 People - German Scripted Monologue Speech Dataset (Smart Car & Smart Home)

351 People - German Scripted Monologue Speech Dataset collected from monologue based on given prompts, covering smart car, smart home, voice assistant domains. Transcribed with text content and other attributes. Our dataset was collected from extensive and diversify speakers(351 native speakers), geographicly speaking, enhancing model performance in real and complex tasks. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.
German speech data German voice data German scripted speech data german scripted monologue speech dataset german smartphone speech data german voice assistant speech data

268 Hours - Arabic(Saudi) Full-Duplex Multi-Channel Customer Service Speech Dataset

268 Hours - Arabic(Saudi) Full-Duplex Multi-Channel Customer Service Speech Dataset. Transcribed with text content, speaker's ID, gender, age and other attributes. Our dataset was collected from extensive and diversify speakers(268 native speakers), geographicly speaking, enhancing model performance in real and complex tasks. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.
full-duplex speech dataset multi-channel audio dataset multi-stream speech dataset Saudi Arabic speech dataset Arabic customer service speech dataset full-duplex multi-channel speech dataset call center Arabic speech data

1100 Hours - Tagalog Full-Duplex Multi-Channel Speech Dataset

1100 Hours - Tagalog Full-Duplex Multi-Channel Speech Dataset, collected from dialogues based on given topics. Transcribed with text content, speaker's ID, gender, age and other attributes. Our dataset was collected from extensive and diversify speakers, geographicly speaking, enhancing model performance in real and complex tasks. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.
multi-stream speech dataset full-duplex dialogue dataset multi-channel audio dataset speech recognition training data multi-channel voice dataset

600 Hours Greek Speech Dataset – Real world Casual Conversation & Monologue for ASR

The 600 Hours Greek Real-World Speech Dataset includes both casual conversations and monologues, covers self-media, conversation, live, variety show and other generic domains, mirroring real-world interactions. Transcribed with text content, speaker's ID, and other attributes. Our dataset was collected from extensive and diversify speakers, geographicly speaking, enhancing model performance in real and complex tasks. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.
greek speech dataset greek ASR training data greek conversation corpus greek monologue speech greek speech recognition dataset speech-to-text greek data greek voice dataset greek transcription dataset

200 Hours – Japanese Spontaneous Dialogue Dataset (Smartphone Multi-Stream Audio)

This dataset consists of 200 hours of multi-stream Japanese spontaneous dialogue recorded via smartphones, featuring natural conversations based on prompted topics. It includes transcriptions with speaker ID, gender, age, and other metadata. Recorded by over 500 native Japanese speakers across diverse regions, this dataset supports advanced speech recognition, speaker diarization, and natural language understanding tasks. It has been validated by major AI companies and fully complies with global privacy regulations including GDPR, CCPA, and PIPL.
Japanese speech dataset spontaneous Japanese dialogue multi-stream Japanese audio data ASR training data Japan smartphone recorded Japanese audio conversational Japanese corpus Japanese voice dataset

600 Hours Norwegian Speech Dataset – Real-world Casual Conversation & Monologue for ASR

The 600 Hours Norwegian Real-World Speech Dataset includes both casual conversations and monologues, covering domains such as self-media, live shows, and other generic domains, mirroring real-world interactions. Transcribed with text content, speaker's ID, and other attributes. Our dataset was collected from extensive and diversify speakers, geographicly speaking, enhancing model performance in real and complex tasks. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.
norwegian speech dataset norwegian ASR training data norwegian conversation corpus norwegian monologue speech norwegian speech recognition dataset speech-to-text norwegian data norwegian voice dataset multilingual speech data norwegian transcription dataset

1,503 Hours – UAE Arabic Speech Dataset for TTS & ASR

This dataset mirrors real-world interactions. Transcribed with text content, speaker's ID, gender, and other attributes. Our dataset was collected from extensive and diversify speakers, geographicly speaking, enhancing model performance in real and complex tasks. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.
UAE Arabic Speech Dataset Arabic speech dataset Arabic Conversational Dataset Arabic Speech Corpus Arabic monologue speech data
. . .

loading

Tailor Your Data Now

Why off-the-shelf Datasets

  • Copyright

    Copyright

    Clear Coyright and Ready to Check
  • Security

    Security

    Properly Authorized Secure to Use
  • Professional

    Professional

    Designed and produced by AI data experts
  • Diversity

    Diversity

    Collected from a varity of real scenes
  • Cost Effective

    Cost Effective

    More Cost-Efficient Than Tailored Data
  • Efficiency

    Efficiency

    Ready-To-Go Deliver in Seconds
f0cb5a7e-c034-4dba-ad84-95d707d68277