en

Please fill in your name

Mobile phone format error

Please enter the telephone

Please enter your company name

Please enter your company email

Please enter the data requirement

Successful submission! Thank you for your support.

Format error, Please fill in again

Confirm

The data requirement cannot be less than 5 words and cannot be pure numbers

m.nexdata.datatang.com

Speech Recognition Datasets

Instantly enhance AI model performance with high quality off-the-shelf datasets.

Language

All
238
Arabic
7
Burmese
3
Chinese Dialects
3
English
45
French
13
German
11
Hindi
7
Indonesian
8
Italian
11
Japanese
12
Korean
14
Malay
4
Mandarin
3
Others
57
Portuguese
14
Russian
6
Spanish
17
Thai
10
Vietnamese
7

Data Type

All
238
Dialogue
117
Read
122

172 Hours - American English Full-Duplex Multi-Channel Speech Dataset

172 Hours - American English Full-Duplex Multi-Channel Speech Dataset, collected from dialogues based on given topics. Transcribed with text content, speaker's ID, gender, age and other attributes. Our dataset was collected from extensive and diversify speakers, geographicly speaking, enhancing model performance in real and complex tasks. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.
American English speech dataset multi-stream speech dataset full-duplex dialogue dataset spontaneous speech dataset smartphone speech data multi-channel audio dataset speech recognition training data dialogue AI dataset

205 Hours - Japanese Full-Duplex Multi-Channel Speech Dataset

205 Hours Japanese Full-Duplex Multi-Channel Speech Dataset is collected from dialogues based on given topics. Transcribed with text content, speaker's ID, gender, age and other attributes. Our dataset was collected from extensive and diversify speakers, geographicly speaking, enhancing model performance in real and complex tasks. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.
Japanese speech dataset spontaneous Japanese dialogue multi-stream Japanese audio data conversational Japanese corpus Japanese voice dataset full-duplex speech dataset multi-stream speech dataset multi-channel audio dataset

200 Hours Korean Full-Duplex Multi-Channel Speech Dataset

This 200 Hours Korean Full-Duplex Multi-Channel Speech Dataset features multi-stream audio recorded via smartphones, simulating natural conversations across a range of everyday topics. Each dialogue is annotated with transcripts, speaker ID, gender, and age. Collected from diverse speakers across various regions in Korea, the dataset enhances AI model robustness in real-world applications. Ideal for training automatic speech recognition (ASR), conversational AI, multilingual dialogue systems, and natural speech processing models. All data complies with GDPR, CCPA, and PIPL privacy standards, ensuring safe and ethical AI training.
Korean speech dataset spontaneous dialogue Korean multi-stream audio dataset conversational Korean speech smartphone-recorded audio dual-speaker dataset real-world Korean conversation full-duplex speech dataset

211 Hours Thai Full-Duplex Speech Dataset for Real-Time Voice Agent Training

This dataset contains 211 hours of Thai full-duplex spontaneous dialogue speech collected from natural conversations based on predefined topics. The dataset includes transcription text, speaker IDs, gender, age information, and other metadata. It was collected from 654 native Thai speakers, enhancing model performance in real and complex tasks. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.
multi-stream speech dataset Thai speech dataset full duplex speech dataset multi speaker speech dataset

351 Speakers German Voice Command Dataset for Smart Car and Smart Home AI

This dataset contains scripted German speech recordings collected from 351 native German speakers based on predefined voice prompts. The recordings cover smart car voice commands, smart home interactions, voice assistant scenarios, and other human-machine interaction applications. Each audio sample is transcribed with corresponding text content and additional metadata attributes. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.
german voice command dataset automotive speech dataset voice assistant training data speech command dataset smart home voice dataset german speech dataset german ASR dataset

268 Hours Arabic Full-Duplex Speech Dataset with Multi-Channel Conversations

This dataset features full-duplex, multi-channel conversations recorded from 268 native Saudi Arabic speakers in authentic customer service scenarios. The dataset includes high-quality audio with verbatim transcriptions and comprehensive metadata, such as speaker ID, gender, age, and other demographic attributes. Our dataset was collected from extensive and diversify speakers(268 native speakers), geographicly speaking, enhancing model performance in real and complex tasks. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.
full-duplex speech dataset multi-channel audio dataset Saudi Arabic speech dataset Arabic customer service speech dataset Arabic speech dataset Saudi Arabic speech dataset Arabic call center dataset

950 Hours - Tagalog Full-Duplex Multi-Channel Speech Dataset

950 Hours - Tagalog Full-Duplex Multi-Channel Speech Dataset, collected from dialogues based on given topics. Transcribed with text content, speaker's ID, gender, age and other attributes. Our dataset was collected from extensive and diversify speakers, geographicly speaking, enhancing model performance in real and complex tasks. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.
multi-stream speech dataset full-duplex dialogue dataset multi-channel audio dataset speech recognition training data multi-channel voice dataset

600 Hours Philippine English Full-Duplex Speech Dataset for Voice Agent Training

This dataset contains 600 hours of Philippine English conversational speech recordings collected from natural dialogues based on predefined topics.The dataset features full-duplex, multi-channel audio recordings. Each audio sample includes accurate transcription and metadata such as speaker ID, gender, age, and other attributes. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.
full duplex speech dataset voice agent dataset conversational AI training data speech to speech dataset multi channel speech dataset Philippine English speech dataset

503 Hours Russian Speech Dataset for Voice AI and Conversational Models

This dataset contains 503 hours of Russian speech recordings collected from native Russian speakers, covering natural conversational and monologue speech scenarios. The dataset captures real-world speaking styles across various topics and daily communication scenarios. Each audio sample includes accurate transcription and speaker metadata, such as speaker ID, gender, and other attributes. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.
russian speech dataset russian ASR dataset russian speech recognition dataset voice AI training data
. . .

loading

Tailor Your Data Now

Why off-the-shelf Datasets

  • Copyright

    Copyright

    Clear Coyright and Ready to Check
  • Security

    Security

    Properly Authorized Secure to Use
  • Professional

    Professional

    Designed and produced by AI data experts
  • Diversity

    Diversity

    Collected from a varity of real scenes
  • Cost Effective

    Cost Effective

    More Cost-Efficient Than Tailored Data
  • Efficiency

    Efficiency

    Ready-To-Go Deliver in Seconds
4e7d5a61-3317-475d-a385-6ad1562c6380