en

Please fill in your name

Mobile phone format error

Please enter the telephone

Please enter your company name

Please enter your company email

Please enter the data requirement

Successful submission! Thank you for your support.

Format error, Please fill in again

Confirm

The data requirement cannot be less than 5 words and cannot be pure numbers

m.nexdata.datatang.com

980K Chinese-Urdu Sentence Pairs – Machine Translation Dataset

Chinese Urdu parallel corpus
Chinese Urdu translation dataset
Chinese Urdu parallel dataset
Chinese Urdu sentence pairs
Chinese Urdu translation corpus

This dataset contains 980,000 Chinese-Urdu parallel sentence pairs stored in TXT format.The data has undergone cleaning, anonymization, and quality inspection to improve data quality and protect sensitive information. The dataset is suitable for machine translation, bilingual language modeling, and translation model training.

Paid Datasets
This is a paid dataset licensed for commercial use. Ready-made datasets are available for immediate integration into AI projects.
SpecificationsSpecifications
Storage format
text
Data content
Chinese-Urdu Parallel Corpus Data
Data size
0.98 million pairs of Chinese-Urdu Parallel Corpus Data. The Chinese sentences contain 19.9 characters on average.
Language
Chinese, Urdu
Accuracy rate
90%
Application scenario
machine translation
Sample Sample
  • 980K Chinese-Urdu Sentence Pairs – Machine Translation Dataset
Recommended DatasetsRecommended Dataset
Tell Us Your Special Needs

Current Project Maturity

Early exploration (no concrete specs yet)
Defined goals, need professional guidance
Active development or optimization phase
Data & labeling experts with clear specifications

By submitting, I agree to the Privacy Protection

f1480a86-a0d0-4f0c-90da-62b2960032ea

94112cd2-2232-4ae2-8778-05b628bf2fb7