en

Please fill in your name

Mobile phone format error

Please enter the telephone

Please enter your company name

Please enter your company email

Please enter the data requirement

Successful submission! Thank you for your support.

Format error, Please fill in again

Confirm

The data requirement cannot be less than 5 words and cannot be pure numbers

500 Hours - German(Germany) Spontaneous Dialogue Smartphone speech dataset

German
audio
data
dataset
conversational
asr data
German data
German dataset
German collection
German asr data
German asr dataset
German asr collection
German discuss data
German discuss dataset
German discuss collection
German discuss asr data
German discuss asr dataset
German discuss asr collection
German small talk data
German small talk dataset
German small talk collection
German small talk asr data
German small talk asr dataset
German small talk asr collection
German conversational data
German conversational dataset
German conversational collection
German conversational asr data
German conversational asr dataset
German conversational asr collection
German chat data
German chat dataset
German chat collection
German chat asr data
German chat asr dataset
German chat asr collection
German communication data
German communication dataset
German communication collection
German communication asr data
German communication asr dataset
German communication asr collection
German speech data
German speech dataset
German speech collection
German speech asr data
German speech asr dataset
German speech asr collection
German talk data
German talk dataset
German talk collection
German talk asr data
German talk asr dataset
German talk asr collection
German conversation data
German conversation dataset
German conversation collection
German conversation asr data
German conversation asr dataset
German conversation asr collection
Germany data
Germany dataset
Germany collection
Germany asr data
Germany asr dataset
Germany asr collection
Germany discuss data
Germany discuss dataset
Germany discuss collection
Germany discuss asr data
Germany discuss asr dataset
Germany discuss asr collection
Germany small talk data
Germany small talk dataset
Germany small talk collection
Germany small talk asr data
Germany small talk asr dataset
Germany small talk asr collection
Germany conversational data
Germany conversational dataset
Germany conversational collection
Germany conversational asr data
Germany conversational asr dataset
Germany conversational asr collection
Germany chat data
Germany chat dataset
Germany chat collection
Germany chat asr data
Germany chat asr dataset
Germany chat asr collection
Germany communication data
Germany communication dataset
Germany communication collection
Germany communication asr data
Germany communication asr dataset
Germany communication asr collection
Germany speech data
Germany speech dataset
Germany speech collection
Germany speech asr data
Germany speech asr dataset
Germany speech asr collection
Germany talk data
Germany talk dataset
Germany talk collection
Germany talk asr data
Germany talk asr dataset
Germany talk asr collection
Germany conversation data
Germany conversation dataset
Germany conversation collection
Germany conversation asr data
Germany conversation asr dataset
Germany conversation asr collection

German(Germany) Spontaneous Dialogue Smartphone speech dataset, collected from dialogues based on given topics. Transcribed with text content, timestamp, speaker's ID, gender and other attributes. Our dataset was collected from extensive and diversify speakers(around 750 native speakers), geographicly speaking, enhancing model performance in real and complex tasks. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.

Paid Datasets
This is a paid datasets for commercial use, research purpose and more. Licensed ready made datasets help jump-start AI projects.
SpecificationsSpecifications
Format
16kHz, 8bit, wav, mono channel;
Content category
Dialogue based on given topics
Recording condition
Low background noise (indoor)
Recording device
Android smartphone, iPhone
Country
Germany(DEU)
Language(Region) Code
de-DE
Language
German
Speaker
Around 750 speakers; male and female balanced;
Features of annotation
Transcription text, timestamp, speaker ID, gender
Accuracy rate
Sentence accuracy rate(SAR) 95%
Sample Sample
  • Audio

    trotzdem irgendwas fühlt, kann man machen.

  • Audio

    Muss ja nicht viel sein. Man kann ja ein bisschen nehmen, dass man nicht zu viel nimmt, aber bis man

  • Audio

    und einfach eine schöne Zeit zu erleben. Man kann ja auch ohne Alkohol zu einer Party gehen

  • Audio

    Das macht es mir auch. Ich gebe dir da vollkommen Recht.

  • Audio

    und trotzdem Spaß haben. Wie siehst denn du das?

Recommended DatasetsRecommended Dataset
501 Hours - Indonesian(Indonesia) Real-world Casual Conversation and Monologue speech dataset

Indonesian(Indonesia) Real-world Casual Conversation and Monologue speech dataset, covers self-media, conversation, live and other generic domains, mirrors real-world interactions. Transcribed with text content, speaker's ID, gender and other attributes. Our dataset was collected from extensive and diversify speakers, geographicly speaking, enhancing model performance in real and complex tasks. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.

Indonesian Colloquial Video text annotation
1,013 Hours - English(Britain) Real-world Casual Conversation and Monologue speech dataset

English(Britain) Real-world Casual Conversation and Monologue speech dataset, covers conversation, self-media, etc, mirrors real-world interactions. Transcribed with text content, speaker's ID, gender, and other attributes. Our dataset was collected from extensive and diversify speakers, geographicly speaking, enhancing model performance in real and complex tasks. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.

Spontaneous Speech british english
157 Hours - Uyghur Spontaneous Dialogue Microphone speech dataset

Uyghur Spontaneous Dialogue Microphone speech dataset, collected from dialogues based on given topics, covering 20+ domains. Transcribed with text content, speaker's ID, gender, age and other attributes. Our dataset was collected from extensive and diversify speakers(326 native speakers), geographicly speaking, enhancing model performance in real and complex tasks. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.

维吾尔语 维语 维吾尔 自然对话 自然对话语音数据 自然对话数据 对话数据集 对话数据 对话语音 对话式AI数据 自然对话语音数据 AI对话语音数据 AI自然语音对话 外语自然对话数据 电话
136 Hours - Korean(Korea) Spontaneous Dialogue Telephony speech dataset

Korean(Korea) Spontaneous Dialogue Telephony speech dataset, collected from dialogues based on given topics, covering 20+ domains. Transcribed with text content, speaker's ID, gender, age and other attributes. Our dataset was collected from extensive and diversify speakers(216 native speakers), geographicly speaking, enhancing model performance in real and complex tasks. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.

Conversational telephone korean
128 Hours - English(Australia) Children Real-world Casual Conversation and Monologue speech dataset

English(Australia) Children Real-world Casual Conversation and Monologue speech dataset, covers self-media, conversation, live, lecture, variety show and other generic domains, mirrors real-world interactions. Transcribed with text content, speaker's ID, gender, age, accent and other attributes. Our dataset was collected from extensive and diversify speakers(12 years old and younger children), geographicly speaking, enhancing model performance in real and complex tasks.rnQuality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.

Australian English Spontaneous Speech text annotation
149 Hours - English(the United Kindom) Children Real-world Casual Conversation and Monologue speech dataset

English(United Kindom) Children Real-world Casual Conversation and Monologue speech dataset, covers self-media, conversation, live, lecture, variety show and other generic domains, mirrors real-world interactions. Transcribed with text content, speaker's ID, gender, age, accent and other attributes. Our dataset was collected from extensive and diversify speakers(12 years old and younger children), geographicly speaking, enhancing model performance in real and complex tasks.rnQuality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.

Spontaneous Speech text annotation British English
145 Hours - Spanish(spain) Children Real-world Casual Conversation and Monologue speech dataset

Spanish(spain) Children Real-world Casual Conversation and Monologue speech dataset, covers self-media, conversation, live, lecture, variety show and other generic domains, mirrors real-world interactions. Transcribed with text content, speaker's ID, gender, age, accent and other attributes. Our dataset was collected from extensive and diversify speakers(12 years old and younger children), geographicly speaking, enhancing model performance in real and complex tasks.rnQuality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.

Spanish Spontaneous Speech text annotation
189 Hours - Spanish(Latin America) Children Real-world Casual Conversation and Monologue speech dataset

Spanish(Latin America) Children Real-world Casual Conversation and Monologue speech dataset, covers self-media, conversation, live, lecture, variety show and other generic domains, mirrors real-world interactions. Transcribed with text content, speaker's ID, gender, age, accent and other attributes. Our dataset was collected from extensive and diversify speakers(12 years old and younger children), geographicly speaking, enhancing model performance in real and complex tasks.rnQuality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.

Latin American Spanish Spontaneous Speech
Tell Us Your Special Needs

By submitting, I agree to the Privacy Protection

31dccc35-ea68-416c-a4f2-6e3cf850d54d

b64026e1-0f8b-492f-9570-ab77f1775b86