Back to catalog
Commercial Speech DatasetAvailable for licensing

Arabic (Saudi Arabia) Conversational Speech

Natural Saudi Arabic conversations from native speakers, capturing regional dialects and real-world conversational speech for ASR, voice AI, and speech model development.

Request a sample
  • Native Speakers
  • Human-Verified
  • Commercial License
  • Production-Ready
Sample Audioar-SA

SPX-AR-SA-0001.wav

Two-party dialogue

0:00--:--
Locale
ar-SA
Audio
16 kHz · 16-bit · Mono
Environment
Quiet indoor
Delivery
WAV · JSONL
Annotation
Human transcript · Utterance-level timestamps
Record Anatomy

One recording. One linked JSONL record.

Each WAV file maps to one JSONL record containing its transcript, speaker attributes, and technical metadata.

Audio Record

SPX-AR-SA-0001.wav

0:0616 kHz · 16-bit · Mono
linked by uniq_id
Linked JSONL Record
manifest.jsonlSPX-AR-SA-0001
{
  "uniq_id": "SPX-AR-SA-0001",
  "duration": 6,
  "language": "Arabic",
  "text": "[Available in licensed delivery]",
  "audio_path": "audio/SPX-AR-SA-0001.wav",
  "spkinfo": {
    "language_code": "ar-SA",
    "spkid": "SPK-014",
    "age_range": "25-34"
  },
  "sample_rate": 16000,
  "bit_depth": 16,
  "channels": 1,
  "environment": "quiet_indoor",
  "dataset_type": "conversational_speech"
}

WAV + Transcript + Metadata

Ready for delivery

Quality Assurance

Validated before delivery.

Every recording is reviewed across audio, transcription, and metadata before delivery.

  1. audio/*.wav

    Audio

    • 16 kHz · 16-bit PCM
    • Signal integrity checked
  2. text

    Transcript

    • Human-reviewed
    • Utterance alignment
  3. manifest.jsonl

    Metadata

    • Schema validated
    • Speaker attributes included
  4. WAV + JSONLReady

    Delivery

    • WAV + JSONL
    • Commercial licensing available

Records that fail validation are excluded from the delivery package.

Checked per record · Human reviewed

Use Cases

From licensed speech to production AI.

Structured audio, transcripts, and metadata support workflows across ASR training, pronunciation modeling, and benchmarking.

This dataset · WAV + JSONL
  • Automatic Speech Recognition

    Train and evaluate ASR models with human-verified reference transcripts.

  • Voice AI & Agents

    Extend voice systems across languages, locales, and accents.

Deployed in production speech systems

Next Step

Evaluate this dataset.

Review representative samples and documentation before licensing.

Request a sample

NDA available for qualified evaluations.