en

Please fill in your name

Mobile phone format error

Please enter the telephone

Please enter your company name

Please enter your company email

Please enter the data requirement

Successful submission! Thank you for your support.

Format error, Please fill in again

Confirm

The data requirement cannot be less than 5 words and cannot be pure numbers

500 Hours - German(Germany) Spontaneous Dialogue Smartphone speech dataset

German
audio
data
dataset
conversational
asr data
German data
German dataset
German collection
German asr data
German asr dataset
German asr collection
German discuss data
German discuss dataset
German discuss collection
German discuss asr data
German discuss asr dataset
German discuss asr collection
German small talk data
German small talk dataset
German small talk collection
German small talk asr data
German small talk asr dataset
German small talk asr collection
German conversational data
German conversational dataset
German conversational collection
German conversational asr data
German conversational asr dataset
German conversational asr collection
German chat data
German chat dataset
German chat collection
German chat asr data
German chat asr dataset
German chat asr collection
German communication data
German communication dataset
German communication collection
German communication asr data
German communication asr dataset
German communication asr collection
German speech data
German speech dataset
German speech collection
German speech asr data
German speech asr dataset
German speech asr collection
German talk data
German talk dataset
German talk collection
German talk asr data
German talk asr dataset
German talk asr collection
German conversation data
German conversation dataset
German conversation collection
German conversation asr data
German conversation asr dataset
German conversation asr collection
Germany data
Germany dataset
Germany collection
Germany asr data
Germany asr dataset
Germany asr collection
Germany discuss data
Germany discuss dataset
Germany discuss collection
Germany discuss asr data
Germany discuss asr dataset
Germany discuss asr collection
Germany small talk data
Germany small talk dataset
Germany small talk collection
Germany small talk asr data
Germany small talk asr dataset
Germany small talk asr collection
Germany conversational data
Germany conversational dataset
Germany conversational collection
Germany conversational asr data
Germany conversational asr dataset
Germany conversational asr collection
Germany chat data
Germany chat dataset
Germany chat collection
Germany chat asr data
Germany chat asr dataset
Germany chat asr collection
Germany communication data
Germany communication dataset
Germany communication collection
Germany communication asr data
Germany communication asr dataset
Germany communication asr collection
Germany speech data
Germany speech dataset
Germany speech collection
Germany speech asr data
Germany speech asr dataset
Germany speech asr collection
Germany talk data
Germany talk dataset
Germany talk collection
Germany talk asr data
Germany talk asr dataset
Germany talk asr collection
Germany conversation data
Germany conversation dataset
Germany conversation collection
Germany conversation asr data
Germany conversation asr dataset
Germany conversation asr collection

German(Germany) Spontaneous Dialogue Smartphone speech dataset, collected from dialogues based on given topics. Transcribed with text content, timestamp, speaker's ID, gender and other attributes. Our dataset was collected from extensive and diversify speakers(around 750 native speakers), geographicly speaking, enhancing model performance in real and complex tasks. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.

Paid Datasets
This is a paid datasets for commercial use, research purpose and more. Licensed ready made datasets help jump-start AI projects.
SpecificationsSpecifications
Format
16kHz, 8bit, wav, mono channel;
Content category
Dialogue based on given topics
Recording condition
Low background noise (indoor)
Recording device
Android smartphone, iPhone
Country
Germany(DEU)
Language(Region) Code
de-DE
Language
German
Speaker
Around 750 speakers; male and female balanced;
Features of annotation
Transcription text, timestamp, speaker ID, gender
Accuracy rate
Sentence accuracy rate(SAR) 95%
Sample Sample
  • Audio

    trotzdem irgendwas fühlt, kann man machen.

  • Audio

    Muss ja nicht viel sein. Man kann ja ein bisschen nehmen, dass man nicht zu viel nimmt, aber bis man

  • Audio

    und einfach eine schöne Zeit zu erleben. Man kann ja auch ohne Alkohol zu einer Party gehen

  • Audio

    Das macht es mir auch. Ich gebe dir da vollkommen Recht.

  • Audio

    und trotzdem Spaß haben. Wie siehst denn du das?

Recommended DatasetsRecommended Dataset
547 Hours - French(France) Spontaneous Dialogue Telephony speech dataset

French(France) Spontaneous Dialogue Telephony speech dataset, collected from dialogues based on given topics, covering 20+ domains. Transcribed with text content, speaker's ID, gender, age and other attributes. Our dataset was collected from extensive and diversify speakers(964 native speakers), geographicly speaking, enhancing model performance in real and complex tasks. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.

Conversational telephone French
499 Hours - Italian(Italy) Spontaneous Dialogue Telephony speech dataset

Italian(Italy) Spontaneous Dialogue Telephony speech dataset, collected from dialogues based on given topics, covering 20+ domains. Transcribed with text content, speaker's ID, gender, age and other attributes. Our dataset was collected from extensive and diversify speakers(676 native speakers), geographicly speaking, enhancing model performance in real and complex tasks. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.

Italian Conversational telephone
1,077 Hours - Thai(Thailand) Spontaneous Dialogue Telephony speech dataset

Thai(Thailand) Spontaneous Dialogue Telephony speech dataset, collected from dialogues based on given topics, covering 20+ domains. Transcribed with text content, speaker's ID, gender, age and other attributes. Our dataset was collected from extensive and diversify speakers(1,986 native speakers), geographicly speaking, enhancing model performance in real and complex tasks. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.

thai Conversational telephone
127 Hours - Portuguese(Brazil) Spontaneous Dialogue Smartphone speech dataset

Portuguese(Brazil) Spontaneous Dialogue Smartphone speech dataset, collected from dialogues based on given topics, covering 20+ domains. Transcribed with text content, speaker's ID, gender, age and other attributes. Our dataset was collected from extensive and diversify speakers(142 native speakers), geographicly speaking, enhancing model performance in real and complex tasks. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.

Conversational speech Portuguese asr data russian asr dataset Brazilian Portuguese
107 Hours - Russian(Russia) Spontaneous Dialogue Smartphone speech dataset

Russian(Russia) Spontaneous Dialogue Smartphone speech dataset, collected from dialogues based on given topics, covering 20+ domains. Transcribed with text content, speaker's ID, gender, age and other attributes. Our dataset was collected from extensive and diversify speakers(134 native speakers), geographicly speaking, enhancing model performance in real and complex tasks. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.

Conversational speech Russian asr data russian asr dataset russia
120 Hours - Burmese(Myanmar) Spontaneous Dialogue Smartphone speech dataset

Burmese(Myanmar) Spontaneous Dialogue Smartphone speech dataset, collected from dialogues based on given topics, covering 20+ domains. Transcribed with text content, speaker's ID, gender, age and other attributes. Our dataset was collected from extensive and diversify speakers(134 native speakers), geographicly speaking, enhancing model performance in real and complex tasks. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.

Conversational speech Burmese asr data Burmese asr dataset
760 Hours - Hindi(India) Spontaneous Dialogue Telephony speech dataset

Hindi(India) Spontaneous Dialogue Telephony speech dataset, collected from dialogues based on given topics, covering 20+ domains. Transcribed with text content, speaker's ID, gender, age and other attributes. Our dataset was collected from extensive and diversify speakers(1,004 native speakers), geographicly speaking, enhancing model performance in real and complex tasks. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.

hindi Conversational Speech phone Hindi discuss data Hindi discuss dataset Hindi discuss collection Hindi small talk data Hindi small talk dataset Hindi small talk collection Hindi conversational data Hindi conversational dataset Hindi conversational collection Hindi chat data Hindi chat dataset Hindi chat collection Hindi communication data Hindi communication dataset Hindi communication collection Hindi speech data Hindi speech dataset Hindi speech collection Hindi talk data Hindi talk dataset Hindi talk collection Hindi conversation data Hindi conversation dataset Hindi conversation collection India discuss data India discuss dataset India discuss collection India small talk data India small talk dataset India small talk collection India conversational data India conversational dataset India conversational collection India chat data India chat dataset India chat collection India communication data India communication dataset India communication collection India speech data India speech dataset India speech collection India talk data India talk dataset India talk collection India conversation data India conversation dataset India conversation collection Indo-Aryan discuss data Indo-Aryan discuss dataset Indo-Aryan discuss collection Indo-Aryan small talk data Indo-Aryan small talk dataset Indo-Aryan small talk collection Indo-Aryan conversational data Indo-Aryan conversational dataset Indo-Aryan conversational collection Indo-Aryan chat data Indo-Aryan chat dataset Indo-Aryan chat collection Indo-Aryan communication data Indo-Aryan communication dataset Indo-Aryan communication collection Indo-Aryan speech data Indo-Aryan speech dataset Indo-Aryan speech collection Indo-Aryan talk data Indo-Aryan talk dataset Indo-Aryan talk collection Indo-Aryan conversation data Indo-Aryan conversation dataset Indo-Aryan conversation collection
749 Hours - Arabic(UAE) Real-world Casual Conversation and Monologue speech dataset

Arabic(UAE) Real-world Casual Conversation and Monologue speech dataset, covers Interview, Speech, Variety, etc, mirrors real-world interactions. Transcribed with text content, speaker's ID, gender, and other attributes. Our dataset was collected from extensive and diversify speakers, geographicly speaking, enhancing model performance in real and complex tasks. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.

Arabic UAE Colloquial video data Arabic Conversation speech data
Tell Us Your Special Needs

By submitting, I agree to the Privacy Protection

b538ec60-93e8-4746-9e1f-1d3daf55344d

345601f5-9107-43b2-bf2d-ce63c041b7be