Back to catalog
Commercial Speech DatasetAvailable for licensing

German Single-Utterance

German single-utterance speech from native speakers, recorded from scripted prompts for ASR training, speech model adaptation, pronunciation modeling, and evaluation.

Request a sample
  • Native Speakers
  • Human-Verified
  • Commercial License
  • Production-Ready
Sample Audiode-DE

SPX-DE-DE-0001.wav

Male · 25–34

0:00--:--
Locale
de-DE
Audio
16 kHz · 16-bit · Mono
Environment
Quiet
Delivery
WAV · JSONL
Annotation
Human transcript · Sentence-level alignment
Record Anatomy

One recording. One linked JSONL record.

Each WAV file maps to one JSONL record containing its transcript, speaker attributes, and technical metadata.

Audio Record

SPX-DE-DE-0001.wav

0:0616 kHz · 16-bit · Mono
linked by uniq_id
Linked JSONL Record
manifest.jsonlSPX-DE-DE-0001
{
  "uniq_id": "SPX-DE-DE-0001",
  "duration": 6,
  "language": "German",
  "text": "[Available in licensed delivery]",
  "audio_path": "audio/SPX-DE-DE-0001.wav",
  "spkinfo": {
    "language_code": "de-DE",
    "spkid": "SPK-712",
    "gender": "male",
    "age_range": "25-34"
  },
  "sample_rate": 16000,
  "bit_depth": 16,
  "channels": 1,
  "environment": "quiet",
  "dataset_type": "single_utterance"
}

WAV + Transcript + Metadata

Ready for delivery

Quality Assurance

Validated before delivery.

Every recording is reviewed across audio, transcription, and metadata before delivery.

  1. audio/*.wav

    Audio

    • 16 kHz · 16-bit PCM
    • Signal integrity checked
  2. text

    Transcript

    • Human-reviewed
    • Sentence-level alignment
  3. manifest.jsonl

    Metadata

    • Schema validated
    • Speaker attributes included
  4. WAV + JSONLReady

    Delivery

    • WAV + JSONL
    • Commercial licensing available

Records that fail validation are excluded from the delivery package.

Checked per record · Human reviewed

Use Cases

From licensed speech to production AI.

Structured audio, transcripts, and metadata support workflows across ASR training, pronunciation modeling, and benchmarking.

This dataset · WAV + JSONL
  • Automatic Speech Recognition

    Train and evaluate ASR models with human-verified reference transcripts.

  • Speech Foundation Models

    Train speech models on structured, multilingual audio data.

  • Voice AI & Agents

    Extend voice systems across languages, locales, and accents.

Deployed in production speech systems

Next Step

Evaluate this dataset.

Review representative samples and documentation before licensing.

Request a sample

NDA available for qualified evaluations.