Datasets ready to license.
12 datasets, 680 hours. Play a sample from any card. Every figure was measured from the audio, and the badges say what each dataset is fit for.
12 of 12 datasets · 680 h of speech
Scripted call-centre conversations in Hindi, insurance, recorded with each speaker on a separate channel. 356 hours. Transcripts are time-aligned and written in Devanagari script, with English words kept as spoken. Layouts across the set: 1,135 two-channel, 710 one side of a call.
356 hTwo channelsTranscribed17% English33.7 dB · clean- ASR
- Full duplex
- Turn-taking
- Voice agents
- Diarisation
- TTS
Hindi speech recognition data: 218,428 single-speaker utterance clips of conversational Hindi, call-centre and everyday, each with its own time-aligned transcript in Devanagari script, English words kept as spoken. 350 hours of audio, 3,377,105 transcribed words.
350 hOne channelTranscribed43.5 dB · some background- ASR
- Full duplex
- Turn-taking
- Voice agents
- Diarisation
- TTS
667 speakersGet a sampleHindi speaker diarization data: whole call-centre and everyday two-speaker conversations with every speaker turn marked, 432,929 turns in RTTM, for training and scoring who spoke when. 321 hours across 1,394 recordings.
321 hOne channelTranscribed9% English30.3 dB · clean- ASR
- Full duplex
- Turn-taking
- Voice agents
- Diarisation
- TTS
Hindi and Tamil full-duplex conversation data: 1,265 two-channel call-centre calls with each speaker on a separate channel, overlaps, backchannels and turn timing preserved, transcripts time-aligned per channel. 264 hours.
264 hTwo channelsTranscribed16% English29.2 dB · some background- ASR
- Full duplex
- Turn-taking
- Voice agents
- Diarisation
- TTS
Tamil speaker diarization data: whole call-centre and everyday two-speaker conversations with every speaker turn marked, 345,825 turns in RTTM, for training and scoring who spoke when. 229 hours across 1,307 recordings.
229 hOne channelTranscribed39% English25.7 dB · some background- ASR
- Full duplex
- Turn-taking
- Voice agents
- Diarisation
- TTS
Tamil speech recognition data: 157,231 single-speaker utterance clips of conversational Tamil, call-centre and everyday, each with its own time-aligned transcript in Tamil script, English words kept as spoken. 201 hours of audio, 1,748,023 transcribed words.
201 hOne channelTranscribed25.6 dB · some background- ASR
- Full duplex
- Turn-taking
- Voice agents
- Diarisation
- TTS
1236 speakersGet a sampleTwo-speaker general conversation in Tamil, recorded on one channel with speakers labelled. 105 hours. Transcripts are time-aligned and written in Tamil script, with English words kept as spoken.
105 hOne channelTranscribed36% English25.7 dB · some background- ASR
- Full duplex
- Turn-taking
- Voice agents
- Diarisation
- TTS
Call-centre conversations in Tamil, consumer surveys, recorded on one channel with speakers labelled. 99 hours. Transcripts are time-aligned and written in Tamil script, with English words kept as spoken.
99 hOne channelTranscribed39% English22.9 dB · some background- ASR
- Full duplex
- Turn-taking
- Voice agents
- Diarisation
- TTS
Two-speaker general conversation in Hindi, recorded on one channel with speakers labelled. 81 hours. Transcripts are time-aligned and written in Devanagari script, with English words kept as spoken.
81 hOne channelTranscribed14% English36.8 dB · clean- ASR
- Full duplex
- Turn-taking
- Voice agents
- Diarisation
- TTS
Scripted call-centre conversations in Tamil, telecom, delivery, e-commerce and banking, recorded with each speaker on a separate channel. 25 hours. Transcripts are time-aligned and written in Tamil script, with English words kept as spoken. Layouts across the set: 2 one side of a call, 132 two-channel.
25 hTwo channelsTranscribed29% English31.2 dB · clean- ASR
- Full duplex
- Turn-taking
- Voice agents
- Diarisation
- TTS
Two-speaker general conversation in Marathi, recorded on one channel with speakers labelled. 9 hours. Transcripts are time-aligned and written in Devanagari script, with English words kept as spoken.
9 hOne channelTranscribed0% English37.8 dB · clean- ASR
- Full duplex
- Turn-taking
- Voice agents
- Diarisation
- TTS
Call-centre conversations in Marathi, banking and insurance, recorded on one channel with speakers labelled. 4 hours. Transcripts are time-aligned and written in Devanagari script, with English words kept as spoken.
4 hOne channelTranscribed0% English35.5 dB · clean- ASR
- Full duplex
- Turn-taking
- Voice agents
- Diarisation
- TTS