[{"@type":"PropertyValue","name":"Data size","value":"451people, 5 images for each person"},{"@type":"PropertyValue","name":"Collection environment","value":"Office, coffee shop, supermarket, apartment"},{"@type":"PropertyValue","name":"Race distribution","value":"151 black people, 150 Caucasians people, 150 brown people ,ranging from teenager to middle-aged people, (Aged between 16 and 60)"},{"@type":"PropertyValue","name":"Gender distribution","value":"226 males, 225 females"},{"@type":"PropertyValue","name":"Data diversity","value":"different poses, different ages, different races, different collection backgrounds"},{"@type":"PropertyValue","name":"Device","value":"computer, cellphone"},{"@type":"PropertyValue","name":"Collecting angles","value":"eye-level angle"},{"@type":"PropertyValue","name":"Data format","value":"the image data format is .jpg, the annotation file (mask) format is .png"},{"@type":"PropertyValue","name":"Annotation content","value":"segmentation annotation of headphones, body, background, glasses"},{"@type":"PropertyValue","name":"Accuracy","value":"based on the accuracy of the actions, the accuracy is more than 97%; Accuracy of semantic segmentation annotation: for each object, the mask edge location errors in x and y directions are less than 5 pixels, and the category label was correctly labeled, which were considered as a qualified annotation; Annotation accuracy: each object is regarded as the unit, annotation accuracy is more than 97%"}]
{"id":1181,"datatype":"1","titleimg":"https://www.nexdata.ai/shujutang/static/image/index/datatang_tuxiang_default.webp","type1":"147","type1str":null,"type2":"149","type2str":null,"dataname":"2,255 Images – Multi-Race Human Body Semantic Segmentation Dataset","datazy":[{"title":"Data size","content":"451people, 5 images for each person","desc":"Data size"},{"title":"Collection environment","content":"Office, coffee shop, supermarket, apartment","desc":"Collection environment"},{"title":"Race distribution","content":"151 black people, 150 Caucasians people, 150 brown people ,ranging from teenager to middle-aged people, (Aged between 16 and 60)","desc":"Race distribution"},{"title":"Gender distribution","content":"226 males, 225 females","desc":"Gender distribution"},{"title":"Data diversity","content":"different poses, different ages, different races, different collection backgrounds","desc":"Data diversity"},{"title":"Device","content":"computer, cellphone","desc":"Device"},{"title":"Collecting angles","content":"eye-level angle","desc":"Collecting angles"},{"title":"Data format","content":"the image data format is .jpg, the annotation file (mask) format is .png","desc":"Data format"},{"title":"Annotation content","content":"segmentation annotation of headphones, body, background, glasses","desc":"Annotation content"},{"title":"Accuracy","content":"based on the accuracy of the actions, the accuracy is more than 97%; Accuracy of semantic segmentation annotation: for each object, the mask edge location errors in x and y directions are less than 5 pixels, and the category label was correctly labeled, which were considered as a qualified annotation; Annotation accuracy: each object is regarded as the unit, annotation accuracy is more than 97%","desc":"Accuracy"}],"datatag":"Human body segmentation,Different poses,Different ages,Different races,Different collection backgrounds","technologydoc":null,"downurl":null,"datainfo":null,"standard":null,"dataylurl":null,"flag":null,"publishtime":null,"createby":null,"createtime":null,"ext1":null,"samplestoreloc":null,"hosturl":null,"datasize":null,"industryPlan":null,"keyInformation":["multiple races","Accuracy of semantic segmentation annotation: for each object, the mask edge location errors in x and y directions are less than 5 pixels, and the category label was correctly labeled, which were considered as a qualified annotation","Annotation accuracy: each object is regarded as the unit, annotation accuracy is more than 97%."],"samplePresentation":[{"name":"/data/apps/damp/temp/ziptemp/APY220628001_demo1712570404821/APY220628001_demo/1.png","url":"https://bj-oss-datatang-03.oss-cn-beijing.aliyuncs.com/filesInfoUpload/data/apps/damp/temp/ziptemp/APY220628001_demo1712570404821/APY220628001_demo/1.png?Expires=4102329599&OSSAccessKeyId=LTAI8NWs2pDolLNH&Signature=P%2F5hYNFsZdPWjVuDcKMzrEkyDuU%3D","intro":"","size":0,"progress":100,"type":"jpg"},{"name":"/data/apps/damp/temp/ziptemp/APY220628001_demo1712570404821/APY220628001_demo/4.png","url":"https://bj-oss-datatang-03.oss-cn-beijing.aliyuncs.com/filesInfoUpload/data/apps/damp/temp/ziptemp/APY220628001_demo1712570404821/APY220628001_demo/4.png?Expires=4102329599&OSSAccessKeyId=LTAI8NWs2pDolLNH&Signature=LaaIVcF4mOBrBNRdUXdcl5m00Yg%3D","intro":"","size":0,"progress":100,"type":"jpg"},{"name":"/data/apps/damp/temp/ziptemp/APY220628001_demo1712570404821/APY220628001_demo/3.png","url":"https://bj-oss-datatang-03.oss-cn-beijing.aliyuncs.com/filesInfoUpload/data/apps/damp/temp/ziptemp/APY220628001_demo1712570404821/APY220628001_demo/3.png?Expires=4102329599&OSSAccessKeyId=LTAI8NWs2pDolLNH&Signature=Gmyem%2BsDGl%2Fh5E%2BmlljFm%2BhCr1M%3D","intro":"","size":0,"progress":100,"type":"jpg"}],"officialSummary":"This dataset includes 2,255 images of 451 people across multiple races. The semantic segmentation area includes headphones, glasses, body and background.This dataset can be used for training AI models in human body segmentation, video conference behavior detection, and smart vision applications.","dataexampl":null,"datakeyword":["human body segmentation dataset","semantic segmentation dataset human","multi-race human dataset","human parsing dataset","video conference dataset","body part segmentation dataset","human behavior recognition dataset","annotated human dataset"],"isDelete":null,"ids":null,"idsList":null,"datasetCode":null,"productStatus":null,"tagTypeEn":"Task Type,Modalities","tagTypeZh":null,"website":null,"samplePresentationList":null,"datazyList":null,"keyInformationList":null,"dataexamplList":null,"bgimg":null,"datazyScriptList":null,"datakeywordListString":null,"sourceShowPage":"computer","dataShowType":"[{\"code\":\"0\",\"language\":\"ZH\"},{\"code\":\"1\",\"language\":\"ZH\"},{\"code\":\"2\",\"language\":\"EN,JP,PT,DE,KO,FR,ES\"},{\"code\":\"4\",\"language\":\"JP\"}]","productNameEn":"451 People –2,255 Images Multi-Races Human Body Semantic Segmentation Data","BGimg":"","voiceBg":["/shujutang/static/image/comm/audio_bg.webp","/shujutang/static/image/comm/audio_bg2.webp","/shujutang/static/image/comm/audio_bg3.webp","/shujutang/static/image/comm/audio_bg4.webp","/shujutang/static/image/comm/audio_bg5.webp"],"firstList":[{"name":"/data/apps/damp/temp/ziptemp/APY220628001_demo1712570404821/APY220628001_demo/5.png","url":"https://bj-oss-datatang-03.oss-cn-beijing.aliyuncs.com/filesInfoUpload/data/apps/damp/temp/ziptemp/APY220628001_demo1712570404821/APY220628001_demo/5.png?Expires=4102329599&OSSAccessKeyId=LTAI8NWs2pDolLNH&Signature=EvCqsOrkGJjU2USkIiZ0Ay4fkm8%3D","intro":"","size":0,"progress":100,"type":"jpg"}]}
https://www.nexdata.ai/shujutang/static/image/index/datatang_tuxiang_default.webp
[{"@type":"ImageObject","embedUrl":"https://bj-oss-datatang-03.oss-cn-beijing.aliyuncs.com/filesInfoUpload/data/apps/damp/temp/ziptemp/APY220628001_demo1712570404821/APY220628001_demo/1.png?Expires=4102329599&OSSAccessKeyId=LTAI8NWs2pDolLNH&Signature=P%2F5hYNFsZdPWjVuDcKMzrEkyDuU%3D"},{"@type":"ImageObject","embedUrl":"https://bj-oss-datatang-03.oss-cn-beijing.aliyuncs.com/filesInfoUpload/data/apps/damp/temp/ziptemp/APY220628001_demo1712570404821/APY220628001_demo/4.png?Expires=4102329599&OSSAccessKeyId=LTAI8NWs2pDolLNH&Signature=LaaIVcF4mOBrBNRdUXdcl5m00Yg%3D"},{"@type":"ImageObject","embedUrl":"https://bj-oss-datatang-03.oss-cn-beijing.aliyuncs.com/filesInfoUpload/data/apps/damp/temp/ziptemp/APY220628001_demo1712570404821/APY220628001_demo/3.png?Expires=4102329599&OSSAccessKeyId=LTAI8NWs2pDolLNH&Signature=Gmyem%2BsDGl%2Fh5E%2BmlljFm%2BhCr1M%3D"},{"@type":"ImageObject","embedUrl":"https://bj-oss-datatang-03.oss-cn-beijing.aliyuncs.com/filesInfoUpload/data/apps/damp/temp/ziptemp/APY220628001_demo1712570404821/APY220628001_demo/5.png?Expires=4102329599&OSSAccessKeyId=LTAI8NWs2pDolLNH&Signature=EvCqsOrkGJjU2USkIiZ0Ay4fkm8%3D"}]
2,255 Images – Multi-Race Human Body Semantic Segmentation Dataset
human body segmentation dataset
semantic segmentation dataset human
multi-race human dataset
human parsing dataset
video conference dataset
body part segmentation dataset
human behavior recognition dataset
annotated human dataset
This dataset includes 2,255 images of 451 people across multiple races. The semantic segmentation area includes headphones, glasses, body and background.This dataset can be used for training AI models in human body segmentation, video conference behavior detection, and smart vision applications.
This is a paid datasets for commercial use, research purpose and more. Licensed ready made datasets help jump-start AI projects.
![Specifications]()
Specifications
Data size
451people, 5 images for each person
Collection environment
Office, coffee shop, supermarket, apartment
Race distribution
151 black people, 150 Caucasians people, 150 brown people ,ranging from teenager to middle-aged people, (Aged between 16 and 60)
Gender distribution
226 males, 225 females
Data diversity
different poses, different ages, different races, different collection backgrounds
Device
computer, cellphone
Collecting angles
eye-level angle
Data format
the image data format is .jpg, the annotation file (mask) format is .png
Annotation content
segmentation annotation of headphones, body, background, glasses
Accuracy
based on the accuracy of the actions, the accuracy is more than 97%; Accuracy of semantic segmentation annotation: for each object, the mask edge location errors in x and y directions are less than 5 pixels, and the category label was correctly labeled, which were considered as a qualified annotation; Annotation accuracy: each object is regarded as the unit, annotation accuracy is more than 97%
![Sample]()
Sample
![Recommended Datasets]()
Recommended Dataset
Tell Us Your Special Needs

What types of computer vision applications can Nexdata’s datasets support?

Nexdata’s computer vision datasets support a wide range of AI applications, including image classification, object detection, image segmentation, facial and human-related recognition, scene understanding, autonomous driving, and other visual perception tasks. Depending on the dataset, data may include images, videos, bounding boxes, polygons, keypoints, segmentation masks, text annotations, and other structured labels.

Can Nexdata customize Computer Vision datasets based on our specific requirements?

Yes. If our off-the-shelf Computer Vision datasets do not fully meet your requirements, Nexdata provides flexible custom data collection, annotation, and curation services. We can customize data based on your target objects, environments, scenarios, camera specifications, geographic locations, data volume, annotation formats, and quality standards to support specific model training and evaluation needs.

How does Nexdata ensure the quality and scalability of its Computer Vision datasets?

Nexdata applies multi-stage quality control throughout data collection, annotation, validation, and delivery. Depending on project requirements, we can implement customized annotation guidelines, multi-level reviews, consistency checks, and quality sampling to ensure dataset accuracy and consistency. Our data collection and processing capabilities can also be scaled to support large-volume Computer Vision projects.
7b4fb553-3105-4162-8724-52df65cc82bc