{"id":1175,"datatype":"1","titleimg":"https://res.datatang.com/asset/productNew/APY220520001.png?Expires=2007353708&OSSAccessKeyId=LTAI5tQwXnJZbubgVfVa1ep9&Signature=0FyQGB%2B3w3JbLzQ2yJRRcO4AaSc%3D","type1":"165","type1str":null,"type2":"166","type2str":null,"dataname":"501 Hours Indian English Accent Speech Dataset with Native Speaker Conversations","datazy":[{"title":"Format","desc":"Format","content":"16kHz, 16 bit, wav, mono channel;"},{"title":"Recording environment","desc":"Recording environment","content":"Low background noise;"},{"title":"Country","desc":"Country","content":"Indian(IND);"},{"title":"Language(Region) Code","desc":"Language(Region) Code","content":"en-IN;"},{"title":"Language","desc":"Language","content":"English;"},{"title":"Features of annotation","desc":"Features of annotation","content":"Transcription text, timestamp, speaker ID, gender."},{"title":"Accuracy Rate","desc":"Accuracy Rate","content":"Sentence Accuracy Rate (SAR) 95%"}],"datatag":"Indian Engliah,Colloquial,Video,Conversation","technologydoc":null,"downurl":null,"datainfo":null,"standard":null,"dataylurl":null,"flag":null,"publishtime":null,"createby":null,"createtime":null,"ext1":null,"samplestoreloc":null,"hosturl":null,"datasize":null,"industryPlan":null,"keyInformation":null,"samplePresentation":[],"officialSummary":"This dataset contains 501 hours of Indian English real-world speech collected from natural conversations and monologue scenarios. The dataset captures authentic Indian English speaking styles across diverse topics and real-life communication scenarios. Each audio sample is transcribed with corresponding text content and enriched with metadata, including speaker ID, gender, and other attributes. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.","dataexampl":null,"datakeyword":["indian english speech dataset","indian english ASR dataset","english accent dataset","conversational speech dataset","speech recognition training data","voice AI training data"],"isDelete":null,"ids":null,"idsList":null,"datasetCode":null,"productStatus":null,"tagTypeEn":"Data Type,Language","tagTypeZh":null,"website":null,"samplePresentationList":null,"datazyList":null,"keyInformationList":null,"dataexamplList":null,"bgimg":null,"datazyScriptList":null,"datakeywordListString":null,"sourceShowPage":"speechRec","dataShowType":"[{\"code\":\"0\",\"language\":\"ZH\"},{\"code\":\"1\",\"language\":\"ZH\"},{\"code\":\"2\",\"language\":\"EN,JP,PT,DE,KO,FR,ES\"},{\"code\":\"3\",\"language\":\"EN\"},{\"code\":\"4\",\"language\":\"JP\"}]","productNameEn":"501 Hours - Indian English Spontaneous Speech Data","BGimg":"brightSpot_audio","voiceBg":["/shujutang/static/image/comm/audio_bg.webp","/shujutang/static/image/comm/audio_bg2.webp","/shujutang/static/image/comm/audio_bg3.webp","/shujutang/static/image/comm/audio_bg4.webp","/shujutang/static/image/comm/audio_bg5.webp"]}
501 Hours Indian English Accent Speech Dataset with Native Speaker Conversations
indian english speech dataset
indian english ASR dataset
english accent dataset
conversational speech dataset
speech recognition training data
voice AI training data
This dataset contains 501 hours of Indian English real-world speech collected from natural conversations and monologue scenarios. The dataset captures authentic Indian English speaking styles across diverse topics and real-life communication scenarios. Each audio sample is transcribed with corresponding text content and enriched with metadata, including speaker ID, gender, and other attributes. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.
This is a paid datasets for commercial use, research purpose and more. Licensed ready made datasets help jump-start AI projects.