268 Hours Arabic Full-Duplex Speech Dataset with Multi-Channel Conversations
This dataset features full-duplex, multi-channel conversations recorded from 268 native Saudi Arabic speakers in authentic customer service scenarios. The dataset includes high-quality audio with verbatim transcriptions and comprehensive metadata, such as speaker ID, gender, age, and other demographic attributes. Our dataset was collected from extensive and diversify speakers(268 native speakers), geographicly speaking, enhancing model performance in real and complex tasks. Quality tested by various AI companies. We strictly adhere to data protection regulations and privacy standards, ensuring the maintenance of user privacy and legal rights throughout the data collection, storage, and usage processes, our datasets are all GDPR, CCPA, PIPL complied.
full-duplex speech dataset multi-channel audio dataset Saudi Arabic speech dataset Arabic customer service speech dataset Arabic speech dataset Saudi Arabic speech dataset Arabic call center dataset