Conversational Audio
Domain Role-Play
Agent-customer role-play built from production-style scripts: 44 scenarios across 9 domains, from card servicing and fraud investigation to flight disruptions, claim filing and salary negotiation. Each conversation follows a scene structure with named speaker roles, so the audio arrives with its intent and context attached.
Specification
- Domains
- 9
- Scenarios
- 44
- Roles
- Named per script, e.g. agent and customer
- Context fields
- script.category, script.sub_category, scene.name
- Channels
- Dual-channel + merged mix
- Language
- English
Delivery layout
- ampllab-speech-to-speech-batch-<timestamp>/
- session_<id>/
- scene_<n>/
- take-<n>/
- <speaker-a>.flac
- <speaker-b>.flac
- merged-<take>.flac
- annotation-<id>.json
Domains
- Banking
- Customer Support
- Travel
- Insurance
- Negotiation
- Sales & Marketing
- Mortgage & Loans
- Identity & Security
- Collections & Hardship
Scenarios (sample)
- Card Servicing
- Fraud Investigation
- Account Opening
- Loan Application
- Flight Disruptions & Refunds
- Hotel Reservations
- Claim Filing
- Policy Renewal
- Technical Support
- Billing Disputes
- Salary Negotiation
- Phishing Scam Reporting
take-<n>/annotation-<id>.jsonIllustrative values
{ "script": { "category": "Banking", "sub_category": "Card Servicing" }, "scene": { "id": "scene_2", "name": "Declined card at checkout" }, "take": { "uid": "7f3c9e21-...", "seq": 4, "duration": 287.42 }, "speakers": { "a41d07c2-...": { "role": "Priya (Card Services Agent)" }, "e9b2f6a8-...": { "role": "Daniel (Customer)" } }, "transcription": [ { "speaker": "e9b2f6a8-...", "start_time": 3.05, "end_time": 7.48, "transcript": "[sigh] Hi, my card got declined twice <overlap>this morning</overlap>.", "emotion": "Frustrated", "languages": "English" }, { "speaker": "a41d07c2-...", "start_time": 7.1, "end_time": 11.86, "transcript": "<overlap>Oh no,</overlap> let me take a look [mm]. Can you confirm the last four digits?", "emotion": "Calm", "languages": "English" } ]}Sample pack
A curated set of takes from this dataset, with audio and annotation files in the delivery format.
More in Conversational Audio
Dual-Channel Conversations
Two-speaker conversations with each voice on its own isolated, time-synchronized channel, plus a merged mix of the take.
179 hours · 369 sessions
Overlapping Speech
Interruptions, back-channels and cross-talk, with every overlap span marked in the transcript. Signal for full-duplex models.
125K overlap spans