AmplLab/samples
Speech data catalog · 2026-10-01

Data Menu

AmplLab records, cleans and annotates conversational speech for speech-to-speech, full-duplex and expressive TTS models. Voice actors perform every conversation on isolated, synchronized channels; trained annotators then transcribe each turn and label emotion, vocal bursts, overlaps and background noise, with automated QA/QC on top.

Accepted audio
179h
Fully annotated
142h
Time-aligned turns
148K
Labelled noise events
49K
Dual-channel take48 kHz
CH1Agent
CH2Customer
calm
[sigh][laughing]
00:0000:0400:0800:12
04.20 → 08.40CH2"[sigh] Hi, my card got declined twice <overlap>this morning</overlap>."
01

Included Metadata

Datasets ship with file, annotation and speaker metadata, produced by AmplLab's recording, cleanup and QA/QC pipeline.

File Metadata

11 fields
  • Raw Audio (FLAC, lossless)
  • Per-speaker Channel Tracks
  • Merged Mixmerged-<take>.flac
  • Sample RatesampleRate
  • Bit DepthbitDepth
  • Codeccodec
  • Channelschannels
  • Durationtake.duration
  • Session / Scene / Take IDsscene.id, take.uid
  • Domain & Scenarioscript.category
  • Speaker Rolesspeakers.*.role

Annotation Metadata

11 fields
  • Speaker-attributed Transcripttranscript
  • Turn Timestampsstart_time, end_time
  • Emotion per Turnemotion
  • Language per Turnlanguages
  • Vocal Bursts & Fillers[laughing] [um]
  • Overlap Spans<overlap>
  • Speaking-style Spans<whisper>
  • Voice Activity Segmentssilero_vad
  • Noise Eventslabel, start, end
  • Speech vs Noise-floor LevelspeakingRMS
  • QA/QC Checkswpm, coverage

Speaker Metadata

7 fields
  • Gendergender
  • Birth Yearbirth_year
  • Native & Secondary Accentnative_accent
  • Primary & Other Languagesprimary_language
  • Voice Descriptionvoice
  • Recording Setuphome_studio
  • Voice-acting Experiencevoice_actor_experience
02

Datasets

Quantities are a snapshot of accepted, annotated data as of 2026-10-01, and collection is ongoing. We also take bespoke requests for data that isn't listed.

Conversational Audio

Expressivity / Emotional Intelligence

Reliability / Voice & Acoustics

Off-menu

Need data that isn't listed?

Tell us the domains, speakers, emotions and labels you need. We script, record, clean and annotate it through the same pipeline, with the same metadata.

Start a request