en

Please fill in your name

Mobile phone format error

Please enter the telephone

Please enter your company name

Please enter your company email

Please enter the data requirement

Successful submission! Thank you for your support.

Format error, Please fill in again

Confirm

The data requirement cannot be less than 5 words and cannot be pure numbers

m.nexdata.datatang.com

1.08 Million English Russian Parallel Corpus Dataset for Machine Translation

english russian parallel corpus
english russian translation dataset
english russian bilingual dataset
parallel corpus dataset
machine translation dataset

This dataset contains 1.08 million English-Russian sentence pairs, the corpus consists of aligned English and Russian text data covering diverse topics and general language usage scenarios. Sensitive content, including political, adult, and personally identifiable information (PII), has been filtered and removed. it can be a base corpus for machine translation, natural language processing, and multilingual AI model development.

Paid Datasets
This is a paid datasets for commercial use, research purpose and more. Licensed ready made datasets help jump-start AI projects.
SpecificationsSpecifications
Storage format
TXT
Data content
English-Russian Parallel Corpus Data
Data size
1.08 million pairs of English-Russian Parallel Corpus Data
Language
English,Russian
Application scenario
machine translation
Sample Sample
  • 1.08 Million English Russian Parallel Corpus Dataset for Machine Translation
Recommended DatasetsRecommended Dataset
Tell Us Your Special Needs

Current Project Maturity

Early exploration (no concrete specs yet)
Defined goals, need professional guidance
Active development or optimization phase
Data & labeling experts with clear specifications

By submitting, I agree to the Privacy Protection

68694888-9a3d-43c0-b9e7-f1e5ccb09519

312fc89e-3077-4635-9a86-1f5ee2d53df1