ASR — 0123 Clean (WAV)
Synthetic clean tone sequence encoding digits 0123 (zero one two three). Pair with the noisy twin and transcript JSON for ASR evaluation.
Short synthetic digit/tone utterances with transcript JSON and clean↔noise pairs — for testing ASR loaders, WER harnesses, and audio preprocessing.
Synthetic clean tone sequence encoding digits 0123 (zero one two three). Pair with the noisy twin and transcript JSON for ASR evaluation.
Noisy twin of the digit sequence 0123 at ~7 dB SNR. Score ASR against the shared transcript JSON.
Ground-truth transcript for the 0123 ASR utterance pair — expected text: “zero one two three”.
Synthetic clean tone sequence encoding digits 4567 (four five six seven). Pair with the noisy twin and transcript JSON for ASR evaluation.
Noisy twin of the digit sequence 4567 at ~7 dB SNR. Score ASR against the shared transcript JSON.
Ground-truth transcript for the 4567 ASR utterance pair — expected text: “four five six seven”.
Synthetic clean tone sequence encoding digits 89 (eight nine). Pair with the noisy twin and transcript JSON for ASR evaluation.
Noisy twin of the digit sequence 89 at ~7 dB SNR. Score ASR against the shared transcript JSON.
Ground-truth transcript for the 89 ASR utterance pair — expected text: “eight nine”.
Synthetic clean tone sequence encoding digits 1357 (one three five seven). Pair with the noisy twin and transcript JSON for ASR evaluation.
Noisy twin of the digit sequence 1357 at ~7 dB SNR. Score ASR against the shared transcript JSON.
Ground-truth transcript for the 1357 ASR utterance pair — expected text: “one three five seven”.
Synthetic clean tone sequence encoding digits 24680 (two four six eight zero). Pair with the noisy twin and transcript JSON for ASR evaluation.
Noisy twin of the digit sequence 24680 at ~7 dB SNR. Score ASR against the shared transcript JSON.
Ground-truth transcript for the 24680 ASR utterance pair — expected text: “two four six eight zero”.
Synthetic clean tone sequence encoding digits 987654 (nine eight seven six five four). Pair with the noisy twin and transcript JSON for ASR evaluation.
Noisy twin of the digit sequence 987654 at ~7 dB SNR. Score ASR against the shared transcript JSON.
Ground-truth transcript for the 987654 ASR utterance pair — expected text: “nine eight seven six five four”.
Short synthetic cue tone for “one” — clean reference for wake-word / digit ASR smoke tests.
Short synthetic cue tone for “zero” — clean reference for wake-word / digit ASR smoke tests.
Extra synthetic digit utterance 1111 (one one one one) — clean twin.
Ground-truth transcript for extra ASR utterance 1111.
Extra synthetic digit utterance 2222 (two two two two) — clean twin.
Ground-truth transcript for extra ASR utterance 2222.
Extra synthetic digit utterance 3333 (three three three three) — clean twin.
Ground-truth transcript for extra ASR utterance 3333.
Extra synthetic digit utterance 4444 (four four four four) — clean twin.
Ground-truth transcript for extra ASR utterance 4444.
Extra synthetic digit utterance 5555 (five five five five) — clean twin.
Ground-truth transcript for extra ASR utterance 5555.
Extra synthetic digit utterance 6666 (six six six six) — clean twin.
Ground-truth transcript for extra ASR utterance 6666.
Extra synthetic digit utterance 7777 (seven seven seven seven) — clean twin.
Ground-truth transcript for extra ASR utterance 7777.
Extra synthetic digit utterance 8888 (eight eight eight eight) — clean twin.
Ground-truth transcript for extra ASR utterance 8888.
Extra synthetic digit utterance 9999 (nine nine nine nine) — clean twin.
Ground-truth transcript for extra ASR utterance 9999.
Extra synthetic digit utterance 0000 (zero zero zero zero) — clean twin.
Ground-truth transcript for extra ASR utterance 0000.
Synthetic DTMF tone sequence for digits “1234”. Pair with the transcript JSON.
Synthetic DTMF tone sequence for digits “567890”. Pair with the transcript JSON.
Ground-truth digit transcript for the DTMF sequence 567890.
Synthetic DTMF tone sequence for digits “*9#”. Pair with the transcript JSON.
Synthetic DTMF tone sequence for digits “042”. Pair with the transcript JSON.
Synthetic DTMF tone sequence for digits “13579”. Pair with the transcript JSON.
JSON Lines ASR training/eval set for the Wave B synthetic digit utterances — each row points at a clean WAV and carries the expected transcript.
We use Google Analytics and show ads via Adsterra and Infolinks. Non-essential cookies and ad scripts run only after you allow the matching categories. See our cookie policy.