[{"@type":"PropertyValue","name":"Format","value":"48,000Hz, 24bit, uncompressed wav, mono channel"},{"@type":"PropertyValue","name":"Recording environment","value":"professional recording studio"},{"@type":"PropertyValue","name":"Recording content","value":"recording corpus reflects the text content of each character"},{"@type":"PropertyValue","name":"Speaker","value":"4 professional CharacterVoice recorded 7 styles, namely criminal subordinate, rough man, little girl, kindly grandma, businessman, grandfather, Non-Commissioned Officer style. Among them, criminal subordinate and rough man are recorded for the same voice actor; The little girl, the grandmother, and the businessman are recorded for the same voice actor. 2 hours/style."},{"@type":"PropertyValue","name":"Annotation","value":"word and pinyin transcription, prosodic boundary annotation"},{"@type":"PropertyValue","name":"Device","value":"microphone"},{"@type":"PropertyValue","name":"Language","value":"Taiwanese mandarin"},{"@type":"PropertyValue","name":"Application scenarios","value":"speech synthesis"}]
{"id":1533,"datatype":"1","titleimg":"https://www.nexdata.ai/shujutang/static/image/index/datatang_yuyin_default.webp","type1":"165","type1str":null,"type2":"219","type2str":null,"dataname":"14 Hours Taiwan Mandarin TTS Dataset – Multi-Style Voices","datazy":[{"title":"Format","desc":"Format","content":"48,000Hz, 24bit, uncompressed wav, mono channel"},{"desc":"Recording environment","content":"professional recording studio","title":"Recording environment"},{"desc":"Recording content","content":"recording corpus reflects the text content of each character","title":"Recording content"},{"desc":"Speaker","content":"4 professional CharacterVoice recorded 7 styles, namely criminal subordinate, rough man, little girl, kindly grandma, businessman, grandfather, Non-Commissioned Officer style. Among them, criminal subordinate and rough man are recorded for the same voice actor; The little girl, the grandmother, and the businessman are recorded for the same voice actor. 2 hours/style.","title":"Speaker"},{"desc":"Annotation","content":"word and pinyin transcription, prosodic boundary annotation","title":"Annotation"},{"desc":"Device","content":"microphone","title":"Device"},{"desc":"Language","content":"Taiwanese mandarin","title":"Language"},{"desc":"Application scenarios","content":"speech synthesis","title":"Application scenarios"}],"datatag":"TTS,Taiwan Mandarin,Multi-style","technologydoc":null,"downurl":null,"datainfo":null,"standard":null,"dataylurl":null,"flag":null,"publishtime":null,"createby":null,"createtime":null,"ext1":null,"samplestoreloc":null,"hosturl":null,"datasize":null,"industryPlan":null,"keyInformation":"","samplePresentation":[],"officialSummary":"This dataset contains 14 hours of Taiwan Mandarin recordings from 4 professional voice actors with 7 speaking styles. The styles are criminal subordinate, rough man, little girl, kind grandma, businessman, grandfather and non-commissioned officer. Professional phonetician participates in the annotation. It is ideal for text-to-speech (TTS), expressive voice generation, virtual avatars, and AI speech synthesis applications.","dataexampl":null,"datakeyword":["Taiwan Mandarin speech dataset","Taiwan Mandarin voice dataset","Taiwan Mandarin speech corpus for AI","Mandarin accent dataset Taiwan","Mandarin TTS dataset"],"isDelete":null,"ids":null,"idsList":null,"datasetCode":null,"productStatus":null,"tagTypeEn":"","tagTypeZh":null,"website":null,"samplePresentationList":null,"datazyList":null,"keyInformationList":null,"dataexamplList":null,"bgimg":null,"datazyScriptList":null,"datakeywordListString":null,"sourceShowPage":"speechSyn","dataShowType":"[{\"code\":\"0\",\"language\":\"ZH\"},{\"code\":\"1\",\"language\":\"ZH\"},{\"code\":\"2\",\"language\":\"EN,JP,PT,DE,KO,FR,ES\"},{\"code\":\"3\",\"language\":\"EN\"},{\"code\":\"4\",\"language\":\"JP\"}]","productNameEn":"14 Hours - Taiwan Mandarin Seven Style Average Tone Speech Synthesis Corpus","BGimg":"brightSpot_audio","voiceBg":["/shujutang/static/image/comm/audio_bg.webp","/shujutang/static/image/comm/audio_bg2.webp","/shujutang/static/image/comm/audio_bg3.webp","/shujutang/static/image/comm/audio_bg4.webp","/shujutang/static/image/comm/audio_bg5.webp"]}
https://www.nexdata.ai/shujutang/static/image/index/datatang_yuyin_default.webp
[]
14 Hours Taiwan Mandarin TTS Dataset – Multi-Style Voices
Taiwan Mandarin speech dataset
Taiwan Mandarin voice dataset
Taiwan Mandarin speech corpus for AI
Mandarin accent dataset Taiwan
Mandarin TTS dataset
This dataset contains 14 hours of Taiwan Mandarin recordings from 4 professional voice actors with 7 speaking styles. The styles are criminal subordinate, rough man, little girl, kind grandma, businessman, grandfather and non-commissioned officer. Professional phonetician participates in the annotation. It is ideal for text-to-speech (TTS), expressive voice generation, virtual avatars, and AI speech synthesis applications.
This is a paid datasets for commercial use, research purpose and more. Licensed ready made datasets help jump-start AI projects.
![Specifications]()
Specifications
Format
48,000Hz, 24bit, uncompressed wav, mono channel
Recording environment
professional recording studio
Recording content
recording corpus reflects the text content of each character
Speaker
4 professional CharacterVoice recorded 7 styles, namely criminal subordinate, rough man, little girl, kindly grandma, businessman, grandfather, Non-Commissioned Officer style. Among them, criminal subordinate and rough man are recorded for the same voice actor; The little girl, the grandmother, and the businessman are recorded for the same voice actor. 2 hours/style.
Annotation
word and pinyin transcription, prosodic boundary annotation
Language
Taiwanese mandarin
Application scenarios
speech synthesis
![Sample]()
Sample
![Recommended Datasets]()
Recommended Dataset
Tell Us Your Special Needs

What languages and voice characteristics are covered by Nexdata’s speech synthesis datasets?

Nexdata offers speech synthesis datasets covering a broad range of languages, dialects, accents, and voice types, supported by extensive global language resources. Our datasets include diverse speakers, speaking styles, emotions, and recording scenarios to support natural and expressive Text-to-Speech (TTS) model development.

Can Nexdata customize speech synthesis datasets for specific languages or requirements?

Yes. If our off-the-shelf TTS datasets do not fully meet your requirements, Nexdata provides flexible custom data collection, transcription, annotation, and quality control services. We can customize datasets based on target languages or dialects, speaker profiles, voice characteristics, emotions, speaking styles, recording environments, and data volume.

How does Nexdata ensure the quality and scalability of speech synthesis datasets?

Nexdata applies multi-stage quality control throughout voice data collection, transcription, annotation, and validation. Combined with our extensive language resources and scalable collection capabilities, we can support both large-scale multilingual TTS projects and specialized datasets for specific voices, accents, emotions, or speech scenarios.
e78448a1-9e9c-4898-a96d-7d620905ca96