[{"@type":"PropertyValue","name":"Storage format","value":"TXT"},{"@type":"PropertyValue","name":"Data content","value":"Chinese-Korean Parallel Corpus Data"},{"@type":"PropertyValue","name":"Data size","value":"12.82 million pairs of Chinese-Korean Parallel Corpus Data. The Chinese sentences contain 25.7 characters on average."},{"@type":"PropertyValue","name":"Language","value":"Chinese, Korean"},{"@type":"PropertyValue","name":"Accuracy rate","value":"90%"},{"@type":"PropertyValue","name":"Application scenario","value":"machine translation"}]
{"id":1200,"datatype":"1","titleimg":"https://www.nexdata.ai/shujutang/static/image/index/datatang_wenben_default.webp","type1":"183","type1str":null,"type2":"185","type2str":null,"dataname":"12.82M Chinese-Korean Sentence Pairs – Parallel Corpus Dataset","datazy":[{"title":"Storage format","content":"TXT","desc":"Storage format"},{"title":"Data content","content":"Chinese-Korean Parallel Corpus Data","desc":"Data content"},{"title":"Data size","content":"12.82 million pairs of Chinese-Korean Parallel Corpus Data. The Chinese sentences contain 25.7 characters on average.","desc":"Data size"},{"title":"Language","content":"Chinese, Korean","desc":"Language"},{"title":"Accuracy rate","content":"90%","desc":"Accuracy rate"},{"title":"Application scenario","content":"machine translation","desc":"Application scenario"}],"datatag":"Chinese,Korean,Chinese-Korean,Parallel Corpus","technologydoc":null,"downurl":null,"datainfo":null,"standard":null,"dataylurl":null,"flag":null,"publishtime":null,"createby":null,"createtime":null,"ext1":null,"samplestoreloc":null,"hosturl":null,"datasize":null,"industryPlan":null,"keyInformation":"","samplePresentation":[{"name":"/data/apps/damp/temp/ziptemp/APY220720005_demo1711015209476/zh-ko ????.png","url":"https://bj-oss-datatang-03.oss-cn-beijing.aliyuncs.com/filesInfoUpload/data/apps/damp/temp/ziptemp/APY220720005_demo1711015209476/zh-ko%20%3F%3F%3F%3F.png?Expires=4102329599&OSSAccessKeyId=LTAI8NWs2pDolLNH&Signature=DiH301E2zIFDQhnNLMtcQ9OQwOs%3D","intro":"","size":0,"progress":100,"type":"jpg"}],"officialSummary":"This dataset contains 12.82 million Chinese-Korean parallel sentence pairs stored in TXT format. The corpus covers multiple domains, including conversational language, travel, news, finance, and other topics.The data has undergone cleaning, anonymization, and quality inspection to improve data quality and protect sensitive information. It can be used as the basic corpus database in the text data files as well as used in machine translation.","dataexampl":null,"datakeyword":["Chinese Korean parallel corpus","Chinese Korean translation dataset","Chinese Korean parallel dataset","Chinese Korean sentence pairs","Chinese Korean translation corpus"],"isDelete":null,"ids":null,"idsList":null,"datasetCode":null,"productStatus":null,"tagTypeEn":"Type","tagTypeZh":null,"website":null,"samplePresentationList":null,"datazyList":null,"keyInformationList":null,"dataexamplList":null,"bgimg":null,"datazyScriptList":null,"datakeywordListString":null,"sourceShowPage":"nlu","dataShowType":"[{\"code\":\"0\",\"language\":\"ZH\"},{\"code\":\"1\",\"language\":\"ZH\"},{\"code\":\"2\",\"language\":\"EN,JP,PT,DE,KO,FR,ES\"},{\"code\":\"3\",\"language\":\"EN\"},{\"code\":\"4\",\"language\":\"JP\"}]","productNameEn":"12,820,000 Groups - Chinese-Korean Parallel Corpus Data","BGimg":"","voiceBg":["/shujutang/static/image/comm/audio_bg.webp","/shujutang/static/image/comm/audio_bg2.webp","/shujutang/static/image/comm/audio_bg3.webp","/shujutang/static/image/comm/audio_bg4.webp","/shujutang/static/image/comm/audio_bg5.webp"]}
https://www.nexdata.ai/shujutang/static/image/index/datatang_wenben_default.webp
[{"@type":"ImageObject","embedUrl":"https://bj-oss-datatang-03.oss-cn-beijing.aliyuncs.com/filesInfoUpload/data/apps/damp/temp/ziptemp/APY220720005_demo1711015209476/zh-ko%20%3F%3F%3F%3F.png?Expires=4102329599&OSSAccessKeyId=LTAI8NWs2pDolLNH&Signature=DiH301E2zIFDQhnNLMtcQ9OQwOs%3D"}]
12.82M Chinese-Korean Sentence Pairs – Parallel Corpus Dataset
Chinese Korean parallel corpus
Chinese Korean translation dataset
Chinese Korean parallel dataset
Chinese Korean sentence pairs
Chinese Korean translation corpus
This dataset contains 12.82 million Chinese-Korean parallel sentence pairs stored in TXT format. The corpus covers multiple domains, including conversational language, travel, news, finance, and other topics.The data has undergone cleaning, anonymization, and quality inspection to improve data quality and protect sensitive information. It can be used as the basic corpus database in the text data files as well as used in machine translation.
This is a paid datasets for commercial use, research purpose and more. Licensed ready made datasets help jump-start AI projects.
![Specifications]()
Specifications
Data content
Chinese-Korean Parallel Corpus Data
Data size
12.82 million pairs of Chinese-Korean Parallel Corpus Data. The Chinese sentences contain 25.7 characters on average.
Application scenario
machine translation
![Sample]()
Sample
![Recommended Datasets]()
Recommended Dataset
Tell Us Your Special Needs
28e7f75e-e49d-43e8-948e-4bb2f52ea6d5