Datasets ready to license.
12 datasets, 680 hours, and 1 sample release in 13 languages. Play a sample from any card. Every figure was measured from the audio, and the badges say what each dataset is fit for.
4 of 13 datasets · 381 h of speech
Phone calls between a person and an AI voice agent in thirteen Indian languages, on recruitment, telecom, healthcare, e-commerce and insurance, each voice on its own channel at 48 kHz. A sample release: 27 calls, 100 minutes, with verbatim transcripts of both sides. Recorded to order in your languages, domains and scenarios, at any volume.
27 calls · 100 min13 languagesTwo channelsTranscribed16% English38.1 dB · clean- ASR
- Full duplex
- Turn-taking
- Voice agents
- Diarisation
- TTS
Scripted call-centre conversations in Hindi, insurance, recorded with each speaker on a separate channel. 356 hours. Transcripts are time-aligned and written in Devanagari script, with English words kept as spoken. Layouts across the set: 1,135 two-channel, 710 one side of a call.
356 hTwo channelsTranscribed17% English33.7 dB · clean- ASR
- Full duplex
- Turn-taking
- Voice agents
- Diarisation
- TTS
Hindi and Tamil full-duplex conversation data: 1,265 two-channel call-centre calls with each speaker on a separate channel, overlaps, backchannels and turn timing preserved, transcripts time-aligned per channel. 264 hours.
264 h2 languagesTwo channelsTranscribed16% English29.2 dB · some background- ASR
- Full duplex
- Turn-taking
- Voice agents
- Diarisation
- TTS
Scripted call-centre conversations in Tamil, telecom, delivery, e-commerce and banking, recorded with each speaker on a separate channel. 25 hours. Transcripts are time-aligned and written in Tamil script, with English words kept as spoken. Layouts across the set: 2 one side of a call, 132 two-channel.
25 hTwo channelsTranscribed29% English31.2 dB · clean- ASR
- Full duplex
- Turn-taking
- Voice agents
- Diarisation
- TTS