Kenpath Labs
Get started

Datasets ready to license.

12 datasets, 680 hours, and 1 sample release in 13 languages. Play a sample from any card. Every figure was measured from the audio, and the badges say what each dataset is fit for.

6 of 13 datasets · 437 h of speech

  • Phone calls between a person and an AI voice agent in thirteen Indian languages, on recruitment, telecom, healthcare, e-commerce and insurance, each voice on its own channel at 48 kHz. A sample release: 27 calls, 100 minutes, with verbatim transcripts of both sides. Recorded to order in your languages, domains and scenarios, at any volume.

    27 calls · 100 min
    13 languages
    Two channels
    Transcribed
    16% English
    38.1 dB · clean
    • ASR
    • Full duplex
    • Turn-taking
    • Voice agents
    • Diarisation
    • TTS
    21 callers · 1 AI voice agentGet a sample
  • Scripted call-centre conversations in Hindi, insurance, recorded with each speaker on a separate channel. 356 hours. Transcripts are time-aligned and written in Devanagari script, with English words kept as spoken. Layouts across the set: 1,135 two-channel, 710 one side of a call.

    356 h
    Two channels
    Transcribed
    17% English
    33.7 dB · clean
    • ASR
    • Full duplex
    • Turn-taking
    • Voice agents
    • Diarisation
    • TTS
    490 speakersGet a sample
  • Hindi speech recognition data: 218,428 single-speaker utterance clips of conversational Hindi, call-centre and everyday, each with its own time-aligned transcript in Devanagari script, English words kept as spoken. 350 hours of audio, 3,377,105 transcribed words.

    350 h
    One channel
    Transcribed
    37.1 dB · some background
    • ASR
    • Full duplex
    • Turn-taking
    • Voice agents
    • Diarisation
    • TTS
    667 speakersGet a sample
  • Hindi speaker diarization data: whole call-centre and everyday two-speaker conversations with every speaker turn marked, 432,929 turns in RTTM, for training and scoring who spoke when. 321 hours across 1,394 recordings.

    321 h
    One channel
    Transcribed
    9% English
    30.3 dB · clean
    • ASR
    • Full duplex
    • Turn-taking
    • Voice agents
    • Diarisation
    • TTS
    619 speakersGet a sample
  • Hindi and Tamil full-duplex conversation data: 1,265 two-channel call-centre calls with each speaker on a separate channel, overlaps, backchannels and turn timing preserved, transcripts time-aligned per channel. 264 hours.

    264 h
    2 languages
    Two channels
    Transcribed
    16% English
    29.2 dB · some background
    • ASR
    • Full duplex
    • Turn-taking
    • Voice agents
    • Diarisation
    • TTS
    529 speakersGet a sample
  • Two-speaker general conversation in Hindi, recorded on one channel with speakers labelled. 81 hours. Transcripts are time-aligned and written in Devanagari script, with English words kept as spoken.

    81 h
    One channel
    Transcribed
    14% English
    36.8 dB · clean
    • ASR
    • Full duplex
    • Turn-taking
    • Voice agents
    • Diarisation
    • TTS
    177 speakersGet a sample