Skip to main content
General Audio / Speech ≥10 hrs/speaker (up to 100)

Customer Support Voice — Empathetic & Measured

Clear, empathetic, measured support-agent studio speech for service-oriented TTS, with English terms embedded in Hindi.

  • Rights-cleared
  • Consent documented
  • 🇮🇳 India data residency
Rights-cleared multilingual dataset catalog illustrating the Customer Support Voice — Empathetic & Measured dataset
Languages
English, Hindi (P0)
Scale
≥10 hrs/speaker (up to 100)
Quality & methodology metrics

Measured, audited, reproducible

Audio: WAV (PCM), mono, 48 kHz (min. 44.1 kHz), 24-bit — studio-recorded (professional studio preferred; treated home studio with prior approval).

Transcripts: both verbatim and normalized, with documented handling of numbers, dates, currency, time, URLs, email, abbreviations, symbols and punctuation, and native code-mixing.

Annotation: emotion, paralinguistics, disfluencies and prosody/emphasis applied pre-delivery by native-speaker annotators. Inter-annotator agreement (Cohen's / Fleiss' κ) is reported per annotation type as our QA standard, with a two-pass QA and adjudication workflow.

Deliverables: audio (WAV); transcripts (TXT/CSV/JSONL); annotations (JSON/JSONL/CSV); metadata (JSON/CSV).

Exclusions warranted in writing: no music/singing, no synthetic or TTS-generated audio, no copyrighted material, no ASR-style/web-scraped data.

Provenance & consent

All audio originates from human speakers recorded in vendor-owned, controlled studio sessions under written contributor agreements that grant commercial reuse, including explicit consent for AI training and commercial voice-cloning / synthetic-voice generation. Never scraped, never repurposed from call-centre or telephony recordings.

Methodology

Studio speech in the customer-support register — empathetic, clear and measured — reflecting real support interactions with code-mixed English terms, entity handling for order IDs, dates and delivery windows.

Data preview

What's inside each record

A representative schema for Customer Support Voice — Empathetic & Measured. The full data card ships the complete field dictionary, value ranges and annotation rubric.

Field Type Example
clip_id string (uuid) "spk_4421…"
audio_path string (wav, 16kHz) "clips/00421.wav"
transcript string (verbatim) "We went to the market this morning…"
duration_sec float 7.42
speaker_meta object {age_band, gender, region}
language string (ISO 639) "eng"
annotator_id string (hashed) "anr_7f3…"
qa_status enum "passed"
consent_ref string "cns_2024_…"

Representative schema — exact fields and value ranges are documented in the dataset card shipped with every licence.

Licensing

License tiers

Choose the tier that matches your use case. Every tier ships with the full data card, provenance log, and quality report.

Non-Exclusive License

Contact us

Worldwide, perpetual, transferable, sublicensable license for commercial TTS training, voice cloning and customer-facing API use. Other parties may also license the same set.

Request access

Category-Exclusive License

Contact us

Exclusive within a defined use-case category (e.g. TTS/voice-cloning) while remaining licensable for other categories. Scoped as an upgrade on the non-exclusive tier.

Talk to sales

Exclusive License

Contact us

Full single-buyer exclusivity — no other party holds or will receive the data — with worldwide, perpetual, transferable, sublicensable rights and provenance transfer.

Talk to sales

Related datasets

See all General datasets

Ready to license Customer Support Voice — Empathetic & Measured?

A senior data PM will scope access, residency, and licensing terms and respond within one business day.