[{"@type":"PropertyValue","name":"Format","value":"16kHz, 16bit, uncompressed wav, mono channel;"},{"@type":"PropertyValue","name":"Recording condition","value":"Low background noise(indoor), without echo;"},{"@type":"PropertyValue","name":"Content category","value":"Generic domain; human-machine interaction; smart home command and in-car command; numbers; news;"},{"@type":"PropertyValue","name":"Recording device","value":"Android Smartphone, iPhone;"},{"@type":"PropertyValue","name":"Speaker","value":"675 speakers totally, with 44% male and 56% female; and 66% speakers of all are in the age group of 18-25,32% speakers of all are in the age group of 26-45, 5% speakers of all are in the age group of 46-60, with a floating rate of 2%;"},{"@type":"PropertyValue","name":"Country","value":"Malaysia(MYS);"},{"@type":"PropertyValue","name":"Language(Region) Code","value":"ms-MY;"},{"@type":"PropertyValue","name":"Language","value":"Malay;"},{"@type":"PropertyValue","name":"Features of annotation","value":"Transcription text;"},{"@type":"PropertyValue","name":"Accuracy Rate","value":"Sentence Accuracy Rate (SAR) 95%"}]
{"id":992,"datatype":"1","titleimg":"https://res.datatang.com/asset/productNew/APY190318004.png?Expires=2007353660&OSSAccessKeyId=LTAI5tQwXnJZbubgVfVa1ep9&Signature=kZ8dDhbctVV1Zln2xbtsA3xdK8Y%3D","type1":"165","type1str":null,"type2":"166","type2str":null,"dataname":"370 Hours Malay Voice AI Dataset for Smart Home and Automotive Speech Recognition","datazy":[{"title":"Format","content":"16kHz, 16bit, uncompressed wav, mono channel;","desc":"Format"},{"title":"Recording condition","content":"Low background noise(indoor), without echo;","desc":"Recording condition"},{"title":"Content category","content":"Generic domain; human-machine interaction; smart home command and in-car command; numbers; news;","desc":"Content category"},{"title":"Recording device","content":"Android Smartphone, iPhone;","desc":"Recording device"},{"title":"Speaker","content":"675 speakers totally, with 44% male and 56% female; and 66% speakers of all are in the age group of 18-25,32% speakers of all are in the age group of 26-45, 5% speakers of all are in the age group of 46-60, with a floating rate of 2%;","desc":"Speaker"},{"title":"Country","content":"Malaysia(MYS);","desc":"Country"},{"title":"Language(Region) Code","content":"ms-MY;","desc":"Language(Region) Code"},{"title":"Language","content":"Malay;","desc":"Language"},{"title":"Features of annotation","content":"Transcription text;","desc":"Features of annotation"},{"title":"Accuracy Rate","content":"Sentence Accuracy Rate (SAR) 95%","desc":"Accuracy Rate"}],"datatag":"Malay,Malaysia,Smartphone,Reading,Scripted Monologue","technologydoc":null,"downurl":null,"datainfo":null,"standard":null,"dataylurl":null,"flag":null,"publishtime":null,"createby":null,"createtime":null,"ext1":null,"samplestoreloc":null,"hosturl":null,"datasize":null,"industryPlan":null,"keyInformation":"","samplePresentation":[{"name":"/data/apps/damp/temp/ziptemp/APY190318004_demo1695808945402/G00273S2197.wav","url":"https://bj-oss-datatang-03.oss-cn-beijing.aliyuncs.com/filesInfoUpload/data/apps/damp/temp/ziptemp/APY190318004_demo1695808945402/G00273S2197.wav?Expires=4102329599&OSSAccessKeyId=LTAI8NWs2pDolLNH&Signature=WRz0h7NEGIwzEosc2btNKjSYpVQ%3D","intro":"Abu Talib pula membalas dengan berkata; \" Jika begitu, anda semua boleh keluar. \".","size":0,"progress":100,"type":"mp3"},{"name":"/data/apps/damp/temp/ziptemp/APY190318004_demo1695808945402/G20366S4429.wav","url":"https://bj-oss-datatang-03.oss-cn-beijing.aliyuncs.com/filesInfoUpload/data/apps/damp/temp/ziptemp/APY190318004_demo1695808945402/G20366S4429.wav?Expires=4102329599&OSSAccessKeyId=LTAI8NWs2pDolLNH&Signature=7vjjgdqwEqztPMqI9PkwR93QNF4%3D","intro":"Turunkan langsir pintar dengan berdasarkan kelembapan dalam bilik.","size":0,"progress":100,"type":"mp3"},{"name":"/data/apps/damp/temp/ziptemp/APY190318004_demo1695808945402/G00238S1001.wav","url":"https://bj-oss-datatang-03.oss-cn-beijing.aliyuncs.com/filesInfoUpload/data/apps/damp/temp/ziptemp/APY190318004_demo1695808945402/G00238S1001.wav?Expires=4102329599&OSSAccessKeyId=LTAI8NWs2pDolLNH&Signature=vNr46%2Be%2BX4WbqAwkXigs%2BMWAvuc%3D","intro":"Terutamanya permainan akhir ini di mana menjadi juara siri adalah pada baris.","size":0,"progress":100,"type":"mp3"},{"name":"/data/apps/damp/temp/ziptemp/APY190318004_demo1695808945402/G00238S5432.wav","url":"https://bj-oss-datatang-03.oss-cn-beijing.aliyuncs.com/filesInfoUpload/data/apps/damp/temp/ziptemp/APY190318004_demo1695808945402/G00238S5432.wav?Expires=4102329599&OSSAccessKeyId=LTAI8NWs2pDolLNH&Signature=uUMcZXNop%2Fh3pvSkRu7Ap54kkBc%3D","intro":"Rendahkan sedikit posisi penyandar kepala.","size":0,"progress":100,"type":"mp3"},{"name":"/data/apps/damp/temp/ziptemp/APY190318004_demo1695808945402/G00273S6449.wav","url":"https://bj-oss-datatang-03.oss-cn-beijing.aliyuncs.com/filesInfoUpload/data/apps/damp/temp/ziptemp/APY190318004_demo1695808945402/G00273S6449.wav?Expires=4102329599&OSSAccessKeyId=LTAI8NWs2pDolLNH&Signature=DVlKYIZHDBrCKRt5BQwoqKn1q%2F0%3D","intro":"Seratus empat puluh enam ribu seratus enam puluh lapan.","size":0,"progress":100,"type":"mp3"}],"officialSummary":"This dataset contains 370 hours of Malay speech recordings collected from 675 native Malay speakers through scripted monologue tasks based on predefined prompts. The dataset covers multiple application scenarios, including general speech, human-machine interaction, smart home voice control, automotive voice commands, numbers, news, and other domains. Each audio sample includes accurate transcription and additional metadata attributes.Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.","dataexampl":null,"datakeyword":["malay speech dataset","malay speech recognition dataset","malay ASR dataset","malay voice assistant dataset","smart home voice dataset"],"isDelete":null,"ids":null,"idsList":null,"datasetCode":null,"productStatus":null,"tagTypeEn":"Data Type,Language","tagTypeZh":null,"website":null,"samplePresentationList":null,"datazyList":null,"keyInformationList":null,"dataexamplList":null,"bgimg":null,"datazyScriptList":null,"datakeywordListString":null,"sourceShowPage":"speechRec","dataShowType":"[{\"code\":\"0\",\"language\":\"ZH\"},{\"code\":\"1\",\"language\":\"ZH\"},{\"code\":\"2\",\"language\":\"EN,JP,PT,DE,KO,FR,ES\"},{\"code\":\"4\",\"language\":\"JP\"}]","productNameEn":"370 Hours - Malay Speech Data by Mobile Phone","BGimg":"brightSpot_audio","voiceBg":["/shujutang/static/image/comm/audio_bg.webp","/shujutang/static/image/comm/audio_bg2.webp","/shujutang/static/image/comm/audio_bg3.webp","/shujutang/static/image/comm/audio_bg4.webp","/shujutang/static/image/comm/audio_bg5.webp"]}
370 Hours Malay Voice AI Dataset for Smart Home and Automotive Speech Recognition
malay speech dataset
malay speech recognition dataset
malay ASR dataset
malay voice assistant dataset
smart home voice dataset
This dataset contains 370 hours of Malay speech recordings collected from 675 native Malay speakers through scripted monologue tasks based on predefined prompts. The dataset covers multiple application scenarios, including general speech, human-machine interaction, smart home voice control, automotive voice commands, numbers, news, and other domains. Each audio sample includes accurate transcription and additional metadata attributes.Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.
This is a paid dataset licensed for commercial use. Ready-made datasets are available for immediate integration into AI projects.
Specifications
Format
16kHz, 16bit, uncompressed wav, mono channel;
Recording condition
Low background noise(indoor), without echo;
Content category
Generic domain; human-machine interaction; smart home command and in-car command; numbers; news;
Recording device
Android Smartphone, iPhone;
Speaker
675 speakers totally, with 44% male and 56% female; and 66% speakers of all are in the age group of 18-25,32% speakers of all are in the age group of 26-45, 5% speakers of all are in the age group of 46-60, with a floating rate of 2%;
Country
Malaysia(MYS);
Language(Region) Code
ms-MY;
Language
Malay;
Features of annotation
Transcription text;
Accuracy Rate
Sentence Accuracy Rate (SAR) 95%
Sample
Audio
Abu Talib pula membalas dengan berkata; " Jika begitu, anda semua boleh keluar. ".
Audio
Turunkan langsir pintar dengan berdasarkan kelembapan dalam bilik.
Audio
Terutamanya permainan akhir ini di mana menjadi juara siri adalah pada baris.
Audio
Rendahkan sedikit posisi penyandar kepala.
Audio
Seratus empat puluh enam ribu seratus enam puluh lapan.
What languages and scenarios are covered by Nexdata’s speech recognition datasets?
Nexdata offers speech recognition datasets covering a broad range of languages, dialects, and accents, backed by extensive global language resources. Our datasets include diverse speakers, acoustic environments, and real-world speech scenarios, supporting multilingual ASR, voice assistants, conversational AI, speech-to-text, and other speech-enabled applications.
Can Nexdata customize speech recognition datasets for specific languages or requirements?
Yes. If our off-the-shelf datasets do not fully meet your requirements, Nexdata provides flexible custom data collection, transcription, and annotation services. We can customize datasets based on target languages or dialects, speaker profiles, recording environments, speech scenarios, data volume, and annotation specifications to meet specific ASR development needs.
How does Nexdata ensure the quality and scalability of speech recognition datasets?
Nexdata applies multi-stage quality control throughout speech data collection, transcription, annotation, and validation. Combined with our extensive language resources and scalable collection capabilities, we can support both large-scale multilingual projects and specialized datasets for specific languages, dialects, accents, and speech scenarios.