en

Please fill in your name

Mobile phone format error

Please enter the telephone

Please enter your company name

Please enter your company email

Please enter the data requirement

Successful submission! Thank you for your support.

Format error, Please fill in again

Confirm

The data requirement cannot be less than 5 words and cannot be pure numbers

m.nexdata.datatang.com

4.72M Chinese-Uyghur Sentence Pairs – Machine Translation Dataset

Chinese Uyghur parallel corpus
Chinese Uyghur translation dataset
Chinese Uyghur sentence pairs
Chinese Uyghur translation corpus

This dataset contains 4.72 million Chinese-Uyghur parallel sentence pairs stored in TXT format. The data has undergone cleaning, anonymization, and quality inspection to improve data quality and protect sensitive information. The dataset is suitable for text data analysis, machine translation and translation model training.

Paid Datasets
This is a paid datasets for commercial use, research purpose and more. Licensed ready made datasets help jump-start AI projects.
SpecificationsSpecifications
Storage format
TXT
Data content
Chinese-Uighur Parallel Corpus Data
Data size
4.72 million pairs of Chinese-Uighur Parallel Corpus Data. The Chinese sentences contain 22 characters on average
Language
Chinese, Uighur
Application scenario
machine translation
Accuracy rate
90%
Sample Sample
  • 4.72M Chinese-Uyghur Sentence Pairs – Machine Translation Dataset
Recommended DatasetsRecommended Dataset
Tell Us Your Special Needs

Current Project Maturity

Early exploration (no concrete specs yet)
Defined goals, need professional guidance
Active development or optimization phase
Data & labeling experts with clear specifications

By submitting, I agree to the Privacy Protection

56b69751-9cd9-4ba7-a93f-bd3e4bedb94f

b53e4c3c-d5e9-43d4-8809-116303dcdf0a