en

Please fill in your name

Mobile phone format error

Please enter the telephone

Please enter your company name

Please enter your company email

Please enter the data requirement

Successful submission! Thank you for your support.

Format error, Please fill in again

Confirm

The data requirement cannot be less than 5 words and cannot be pure numbers

m.nexdata.datatang.com

7.29M Chinese-Vietnamese Sentence Pairs – Machine Translation Dataset

Chinese Vietnamese parallel corpus
Chinese Vietnamese translation dataset
Chinese Vietnamese parallel dataset
Chinese Vietnamese sentence pairs
Chinese Vietnamese translation corpus

This dataset contains 7.29 million Chinese-Vietnamese parallel sentence pairs stored in TXT format. The corpus covers multiple domains, including tourism, healthcare, daily life, news, and other topics. The data has undergone cleaning, anonymization, and quality inspection to improve data quality and protect sensitive information. It can be used as a basic corpus for text data analysis in fields such as machine translation.

Paid Datasets
This is a paid datasets for commercial use, research purpose and more. Licensed ready made datasets help jump-start AI projects.
SpecificationsSpecifications
Storage format
TXT
Data content
Chinese-Vietnamese Parallel Corpus Data
Data size
7.29 million pairs of Chinese-Vietnamese Parallel Corpus Data
Language
Chinese,Vietnamese
Application scenario
machine translation
Accuracy rate
90%
Sample Sample
  • 7.29M Chinese-Vietnamese Sentence Pairs – Machine Translation Dataset
Recommended DatasetsRecommended Dataset
Tell Us Your Special Needs

Current Project Maturity

Early exploration (no concrete specs yet)
Defined goals, need professional guidance
Active development or optimization phase
Data & labeling experts with clear specifications

By submitting, I agree to the Privacy Protection

f2b7ceb8-2e3f-4c09-9380-b1c0f1dfbe72

a9cd0421-fe6b-411c-8907-8733f35148c2