Skip to content
Testaroo

Explore the test library

Find files, editable templates and browser test targets by what you need to make or test. The directory below is cut by format; the two collections under it cut the same library by subject and by workflow.

396 results

Page 14 of 17; 24 results per page.

Show the canonical directory
Preview of Reservation voicemail: codec mp3 32
mp3
20 KB
Actual file preview for Reservation voicemail: codec mp3 32

Reservation voicemail: codec mp3 32

A 5 second excerpt of this group's take, codec. Encoded and decoded through a real speech codec, so the artefacts are the ones the format actually produces rather than a simulation of them. Delivered as mp3 at 24000 Hz mono. Scored against plate.flac in the same group, which is the same five seconds undamaged. The parameters in the specification are what the file MEASURED, not what was requested of it.

File
MP3 · Speech Ladders · mp3
Use case
ASR testingAudio analysis+1· Conversion set
Preview of Reservation voicemail: codec mulaw 8k
wav
39.2 KB
Actual file preview for Reservation voicemail: codec mulaw 8k

Reservation voicemail: codec mulaw 8k

A 5 second excerpt of this group's take, codec. Encoded and decoded through a real speech codec, so the artefacts are the ones the format actually produces rather than a simulation of them. Delivered as pcm_mulaw at 8000 Hz mono. Scored against plate.flac in the same group, which is the same five seconds undamaged. The parameters in the specification are what the file MEASURED, not what was requested of it.

File
WAV · Speech Ladders · pcm_mulaw
Use case
ASR testingAudio analysis+1· Conversion set
Preview of Reservation voicemail: codec opus 24
opus
14.8 KB
Actual file preview for Reservation voicemail: codec opus 24

Reservation voicemail: codec opus 24

A 5 second excerpt of this group's take, codec. Encoded and decoded through a real speech codec, so the artefacts are the ones the format actually produces rather than a simulation of them. Delivered as opus at 48000 Hz mono. Scored against plate.flac in the same group, which is the same five seconds undamaged. The parameters in the specification are what the file MEASURED, not what was requested of it.

File
OPUS · Speech Ladders · opus
Use case
ASR testingAudio analysis+1· Conversion set
Preview of Reservation voicemail: dropouts
flac
141.9 KB
Actual file preview for Reservation voicemail: dropouts

Reservation voicemail: dropouts

A 5 second excerpt of this group's take, dropouts. Short spans silenced outright, as a packet-switched call drops them, so a transcriber must decide between a pause and a missing word. Delivered as flac at 24000 Hz mono. Scored against plate.flac in the same group, which is the same five seconds undamaged. The parameters in the specification are what the file MEASURED, not what was requested of it.

File
FLAC · Speech Ladders · flac
Use case
ASR testingAudio analysis+1· Conversion set
Preview of Reservation voicemail: full spoken take
flac
455.7 KB
Actual file preview for Reservation voicemail: full spoken take

Reservation voicemail: full spoken take

The complete 15.6 second utterance at 24000 Hz mono FLAC, spoken by the bf_emma voice at 1.0x rate. The five-second excerpt in this group's ladder is cut from this take, so anything needing real duration uses this file and anything comparing damage uses the ladder. Lossless, so it is the archival reference for the whole group.

File
FLAC · Speech Ladders · flac
Use case
ASR testingAudio analysis· Conversion set
Preview of Reservation voicemail: noise snr0
flac
329.3 KB
Actual file preview for Reservation voicemail: noise snr0

Reservation voicemail: noise snr0

A 5 second excerpt of this group's take, noise. Broadband noise mixed in at a measured signal-to-noise ratio, which is where a denoiser is either working or not. Delivered as flac at 24000 Hz mono. Scored against plate.flac in the same group, which is the same five seconds undamaged. The parameters in the specification are what the file MEASURED, not what was requested of it.

File
FLAC · Speech Ladders · flac
Use case
ASR testingAudio analysis+1· Conversion set
Preview of Reservation voicemail: noise snr10
flac
310.4 KB
Actual file preview for Reservation voicemail: noise snr10

Reservation voicemail: noise snr10

A 5 second excerpt of this group's take, noise. Broadband noise mixed in at a measured signal-to-noise ratio, which is where a denoiser is either working or not. Delivered as flac at 24000 Hz mono. Scored against plate.flac in the same group, which is the same five seconds undamaged. The parameters in the specification are what the file MEASURED, not what was requested of it.

File
FLAC · Speech Ladders · flac
Use case
ASR testingAudio analysis+1· Conversion set
Preview of Reservation voicemail: noise snr20
flac
295.9 KB
Actual file preview for Reservation voicemail: noise snr20

Reservation voicemail: noise snr20

A 5 second excerpt of this group's take, noise. Broadband noise mixed in at a measured signal-to-noise ratio, which is where a denoiser is either working or not. Delivered as flac at 24000 Hz mono. Scored against plate.flac in the same group, which is the same five seconds undamaged. The parameters in the specification are what the file MEASURED, not what was requested of it.

File
FLAC · Speech Ladders · flac
Use case
ASR testingAudio analysis+1· Conversion set
Preview of Reservation voicemail: plate
flac
151.9 KB
Actual file preview for Reservation voicemail: plate

Reservation voicemail: plate

A 5 second excerpt of this group's take, reference. The clean five-second excerpt every other file in this ladder is scored against. Delivered as flac at 24000 Hz mono. The parameters in the specification are what the file MEASURED, not what was requested of it.

File
FLAC · Speech Ladders · flac
Use case
ASR testingAudio analysis+1· Conversion set
Preview of Reservation voicemail: rate 16000
flac
187 KB
Actual file preview for Reservation voicemail: rate 16000

Reservation voicemail: rate 16000

A 5 second excerpt of this group's take, resample. The same audio at a different sample rate, which is the step most speech pipelines get wrong silently by assuming their model's rate. Delivered as flac at 16000 Hz mono. Scored against plate.flac in the same group, which is the same five seconds undamaged. The parameters in the specification are what the file MEASURED, not what was requested of it.

File
FLAC · Speech Ladders · flac
Use case
ASR testingAudio analysis+1· Conversion set
Preview of Reservation voicemail: rate 44100
flac
400.8 KB
Actual file preview for Reservation voicemail: rate 44100

Reservation voicemail: rate 44100

A 5 second excerpt of this group's take, resample. The same audio at a different sample rate, which is the step most speech pipelines get wrong silently by assuming their model's rate. Delivered as flac at 44100 Hz mono. Scored against plate.flac in the same group, which is the same five seconds undamaged. The parameters in the specification are what the file MEASURED, not what was requested of it.

File
FLAC · Speech Ladders · flac
Use case
ASR testingAudio analysis+1· Conversion set
Preview of Reservation voicemail: rate 8000
flac
96.9 KB
Actual file preview for Reservation voicemail: rate 8000

Reservation voicemail: rate 8000

A 5 second excerpt of this group's take, resample. The same audio at a different sample rate, which is the step most speech pipelines get wrong silently by assuming their model's rate. Delivered as flac at 8000 Hz mono. Scored against plate.flac in the same group, which is the same five seconds undamaged. The parameters in the specification are what the file MEASURED, not what was requested of it.

File
FLAC · Speech Ladders · flac
Use case
ASR testingAudio analysis+1· Conversion set
Preview of Reservation voicemail: recogniser transcript
txt
214 B
Actual file preview for Reservation voicemail: recogniser transcript

Reservation voicemail: recogniser transcript

What a speech recogniser returned for this take. It normalises: spoken "forty three dollars and eighteen cents" comes back as digits and a currency symbol. Compared with the reference script without normalising first, this disagrees on every number in the utterance and a word error rate computed that way reports a fault that does not exist. That gap is the point of the pair.

File
TXT · Speech Ladders
Use case
ASR testing· Paired fixture
Preview of Reservation voicemail: reference script
txt
262 B
Actual file preview for Reservation voicemail: reference script

Reservation voicemail: reference script

Exactly what was spoken, which is what the synthesiser was given: numbers, currency and times are spelled out as words because that is how they were said. This is ground truth by construction - it existed before the audio did - and it is deliberately NOT the same string as the recogniser transcript paired with it.

File
TXT · Speech Ladders
Use case
ASR testing· Paired fixture
Preview of Reservation voicemail: timed captions
srt
379 B
Actual file preview for Reservation voicemail: timed captions

Reservation voicemail: timed captions

The recogniser's output with timings, 5 cues over the full take. Carries the same normalisation as the plain transcript in this group, so it inherits the same scoring trap, and adds the timing dimension: a cue that starts before the word is spoken is a different defect from a cue with the wrong words in it.

File
SRT · Speech Ladders · 5 cues
Use case
ASR testingMedia accessibility· Conversion set
Preview of Silence-in-Middle: Asym Lead
wav
46.9 KB
Actual file preview for Silence-in-Middle: Asym Lead

Silence-in-Middle: Asym Lead

440 Hz tone with 0.5s of digital silence in the middle (0.2s + silence + 0.8s). Fixture for mid-gap trim / split tools.

File
WAV · Silence Middle · 16000
Use case
Auto-trim testingAudio analysis· Conversion set
Preview of Silence-in-Middle: Asym Trail
wav
46.9 KB
Actual file preview for Silence-in-Middle: Asym Trail

Silence-in-Middle: Asym Trail

440 Hz tone with 0.5s of digital silence in the middle (0.8s + silence + 0.2s). Fixture for mid-gap trim / split tools.

File
WAV · Silence Middle · 16000
Use case
Auto-trim testingAudio analysis· Conversion set
Preview of Silence-in-Middle: Long Gap
wav
71.9 KB
Actual file preview for Silence-in-Middle: Long Gap

Silence-in-Middle: Long Gap

440 Hz tone with 1.5s of digital silence in the middle (0.4s + silence + 0.4s). Fixture for mid-gap trim / split tools.

File
WAV · Silence Middle · 16000
Use case
Auto-trim testingAudio analysis· Conversion set