{"id":1846,"datatype":"1","titleimg":"https://www.nexdata.ai/shujutang/static/image/index/datatang_yuyin_default.webp","type1":"165","type1str":null,"type2":"166","type2str":null,"dataname":"632 Hours Danish Conversational Speech Dataset with Transcriptions","datazy":[{"title":"Format","content":"16k Hz, 16 bit, wav, mono channel;"},{"title":"Recording environment","content":"Low background noise;"},{"title":"Country","content":"Denmark;"},{"title":"Language","content":"Danish;"},{"title":"Features of annotation","content":"Transcription text, timestamp, speaker ID, gender, noise."},{"title":"Accuracy Rate","content":"Word Accuracy Rate (WAR) 98%"}],"datatag":"danish,asr","technologydoc":null,"downurl":null,"datainfo":null,"standard":null,"dataylurl":null,"flag":null,"publishtime":null,"createby":null,"createtime":null,"ext1":null,"samplestoreloc":null,"hosturl":null,"datasize":null,"industryPlan":null,"keyInformation":null,"samplePresentation":[],"officialSummary":"This Danish speech dataset features real-world casual conversations and monologues, reflecting authentic everyday interactions. The dataset includes high-quality audio recordings with transcriptions, speaker IDs, gender information, and other relevant metadata. Our dataset was collected from speakers with diverse geographical and background profiles, thereby enhancing the model's performance in real-world, complex tasks; the dataset has undergone quality validation by multiple AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.","dataexampl":null,"datakeyword":["danish speech dataset","danish conversational speech dataset","danish speech corpus","danish asr dataset","danish conversation dataset","danish dialogue dataset","nordic speech dataset"],"isDelete":null,"ids":null,"idsList":null,"datasetCode":null,"productStatus":null,"tagTypeEn":"Data Type,Language","tagTypeZh":null,"website":null,"samplePresentationList":null,"datazyList":null,"keyInformationList":null,"dataexamplList":null,"bgimg":null,"datazyScriptList":null,"datakeywordListString":null,"sourceShowPage":"speechRec","dataShowType":"[{\"code\":\"0\",\"language\":\"ZH\"},{\"code\":\"1\",\"language\":\"ZH\"},{\"code\":\"2\",\"language\":\"EN\"},{\"code\":\"3\",\"language\":\"EN\"}]","productNameEn":"632 Hours - Danish Real-world Casual Conversation and Monologue speech dataset","BGimg":"brightSpot_audio","voiceBg":["/shujutang/static/image/comm/audio_bg.webp","/shujutang/static/image/comm/audio_bg2.webp","/shujutang/static/image/comm/audio_bg3.webp","/shujutang/static/image/comm/audio_bg4.webp","/shujutang/static/image/comm/audio_bg5.webp"]}
632 Hours Danish Conversational Speech Dataset with Transcriptions
danish speech dataset
danish conversational speech dataset
danish speech corpus
danish asr dataset
danish conversation dataset
danish dialogue dataset
nordic speech dataset
This Danish speech dataset features real-world casual conversations and monologues, reflecting authentic everyday interactions. The dataset includes high-quality audio recordings with transcriptions, speaker IDs, gender information, and other relevant metadata. Our dataset was collected from speakers with diverse geographical and background profiles, thereby enhancing the model's performance in real-world, complex tasks; the dataset has undergone quality validation by multiple AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.
This is a paid datasets for commercial use, research purpose and more. Licensed ready made datasets help jump-start AI projects.