Dataset Catalog

Explore the speech dataset catalog.

Search and filter commercially licensed multilingual speech datasets for ASR, speech foundation models, voice agents, and more.

60 datasets · Page 1 of 5

Arabic (Bahrain) Spontaneous Dialogue

Spontaneous Bahraini Arabic conversations from native speakers, capturing Gulf pronunciation and natural dialogue patterns for ASR, voice agents, and multilingual speech systems.

Locale
ar-BH
Dataset Type
Spontaneous Dialogue
Audio Format
16 kHz / 16-bit PCM WAV
Domain
General

Arabic (Kuwait) Spontaneous Dialogue

Native Kuwaiti Arabic dialogue capturing conversational rhythm, regional pronunciation, and spontaneous speaking styles for production speech recognition and conversational AI.

Locale
ar-KW
Dataset Type
Spontaneous Dialogue
Audio Format
16 kHz / 16-bit PCM WAV
Domain
General

Arabic (MSA) Scripted Monologue

Modern Standard Arabic scripted monologues recorded by native speakers for controlled ASR training, speech model adaptation, and pronunciation-focused evaluation.

Locale
ar
Dataset Type
Scripted Monologue
Audio Format
16 kHz / 16-bit PCM WAV
Domain
General

Arabic (MSA) Spontaneous Dialogue

Spontaneous Modern Standard Arabic conversations designed to capture natural dialogue patterns for ASR, conversational AI, and multilingual speech model development.

Locale
ar
Dataset Type
Spontaneous Dialogue
Audio Format
16 kHz / 16-bit PCM WAV
Domain
General

Arabic (Oman) Spontaneous Dialogue

Omani Arabic conversational recordings reflecting regional pronunciation and spontaneous spoken interaction for multilingual ASR and enterprise voice applications.

Locale
ar-OM
Dataset Type
Spontaneous Dialogue
Audio Format
16 kHz / 16-bit PCM WAV
Domain
General

Arabic (Qatar) Spontaneous Dialogue

Spontaneous Qatari Arabic dialogue with native pronunciation and natural conversational flow for ASR, voice agents, and speech foundation model training.

Locale
ar-QA
Dataset Type
Spontaneous Dialogue
Audio Format
16 kHz / 16-bit PCM WAV
Domain
General

Arabic (Saudi Arabia) Spontaneous Dialogue

Natural Saudi Arabic conversations from native speakers, designed for ASR, automotive voice systems, voice agents, and multilingual speech model development.

Locale
ar-SA
Dataset Type
Spontaneous Dialogue
Audio Format
16 kHz / 16-bit PCM WAV
Domain
General

Arabic (UAE) Spontaneous Dialogue

Emirati Arabic conversational speech from native speakers, supporting voice assistants, speech foundation models, automotive systems, and multilingual ASR.

Locale
ar-AE
Dataset Type
Spontaneous Dialogue
Audio Format
16 kHz / 16-bit PCM WAV
Domain
General

Burmese Telephony Speech

Burmese telephony conversations recorded from native speakers, supporting ASR, IVR, speech analytics, and voice systems operating under telephone-channel conditions.

Locale
my-MM
Dataset Type
Telephony Speech
Audio Format
16 kHz / 16-bit PCM WAV
Domain
General

Cantonese Spontaneous Dialogue

Native Cantonese spontaneous conversations capturing natural speaking rhythm and dialogue structure for ASR, voice agents, and multilingual speech applications.

Locale
yue
Dataset Type
Spontaneous Dialogue
Audio Format
16 kHz / 16-bit PCM WAV
Domain
General

Czech Spontaneous IVR

Spontaneous Czech IVR responses recorded by native speakers for voice interface development, intent recognition, ASR, and automated customer interaction systems.

Locale
cs-CZ
Dataset Type
Spontaneous IVR
Audio Format
16 kHz / 16-bit PCM WAV
Domain
General

Danish Scripted Monologue

Native Danish scripted monologues providing controlled speech recordings for ASR training, pronunciation modeling, speech model adaptation, and evaluation.

Locale
da-DK
Dataset Type
Scripted Monologue
Audio Format
16 kHz / 16-bit PCM WAV
Domain
General