en

Please fill in your name

Mobile phone format error

Please enter the telephone

Please enter your company name

Please enter your company email

Please enter the data requirement

Successful submission! Thank you for your support.

Format error, Please fill in again

Confirm

The data requirement cannot be less than 5 words and cannot be pure numbers

m.nexdata.datatang.com

900-person multi-ethnic facial video collection dataset

Multi-ethnic
voice
face

900-person multi-ethnic facial video collection dataset,Each participant is associated with two video clips and one TXT document.One video captures the participant reading the content of the TXT document aloud, while maintaining natural lip movements and facial expressions.The other video captures the participant remaining completely silent throughout, with a natural facial expression.The participants are instructed to simulate a videoconference scenario: they look straight at the camera and frame their upper body (from the waist/chest upward).Acquisition diversity: Multi-ethnic, multilingual, multi-topic document content, mix of landscape and portrait orientations, multiple backgrounds, multiple age groupsLabel accuracy exceeds 95%; the matching accuracy between the video's read-aloud audio and the document sentences is greater than 95%.The data can be used for application scenarios such as cross-ethnic face recognition, liveness detection, digital humans, and video face swapping.

Paid Datasets
This is a paid datasets for commercial use, research purpose and more. Licensed ready made datasets help jump-start AI projects.
SpecificationsSpecifications
Data Content
Each participant is associated with two video clips and one TXT document.One video captures the participant reading the content of the TXT document aloud, while maintaining natural lip movements and facial expressions.The other video captures the participant remaining completely silent throughout, with a natural facial expression.The participants are instructed to simulate a video-conference scenario: they look straight at the camera and frame their upper body (from the waist/chest upward).
Data Scale
900 persons
Gender distribution
Male, Female
Ethnic distribution
Asian, White, Black
Acquisition environment
Indoor
Acquisition devices
Smartphones, webcams (builtin laptop cameras or external USB cameras)
Acquisition diversity
Multi-ethnic, multilingual, multi-topic document content, mix of landscape and portrait orientations, multiple backgrounds, multiple age groups
Data format
MP4, MOV, and other video formats
Video pixel count
The total number of pixels per video is between 777,600 and 8,294,400
Accuracy
Label accuracy exceeds 95%; the matching accuracy between the video's read-aloud audio and the document sentences is greater than 95%.
Sample Sample
  • 900-person multi-ethnic facial video collection dataset
  • 900-person multi-ethnic facial video collection dataset
  • 900-person multi-ethnic facial video collection dataset
Recommended DatasetsRecommended Dataset
Tell Us Your Special Needs

Current Project Maturity

Early exploration (no concrete specs yet)
Defined goals, need professional guidance
Active development or optimization phase
Data & labeling experts with clear specifications

By submitting, I agree to the Privacy Protection

What types of computer vision applications can Nexdata’s datasets support?

Nexdata’s computer vision datasets support a wide range of AI applications, including image classification, object detection, image segmentation, facial and human-related recognition, scene understanding, autonomous driving, and other visual perception tasks. Depending on the dataset, data may include images, videos, bounding boxes, polygons, keypoints, segmentation masks, text annotations, and other structured labels.

Can Nexdata customize Computer Vision datasets based on our specific requirements?

Yes. If our off-the-shelf Computer Vision datasets do not fully meet your requirements, Nexdata provides flexible custom data collection, annotation, and curation services. We can customize data based on your target objects, environments, scenarios, camera specifications, geographic locations, data volume, annotation formats, and quality standards to support specific model training and evaluation needs.

How does Nexdata ensure the quality and scalability of its Computer Vision datasets?

Nexdata applies multi-stage quality control throughout data collection, annotation, validation, and delivery. Depending on project requirements, we can implement customized annotation guidelines, multi-level reviews, consistency checks, and quality sampling to ensure dataset accuracy and consistency. Our data collection and processing capabilities can also be scaled to support large-volume Computer Vision projects.

494a8d50-f54c-43a5-b331-7e15d7a537d3

2232c383-3927-4fca-aada-098ab2a9e2f0