Team Ai
Datasetpublic

eQOURSE/multilingual-speech

Multilingual Indian Conversational Speech A dataset of naturalistic, spontaneous two-speaker conversations across 13 Indian languages, with segment-level transcripts, speaker profiles, timestamps, and recording metadata. Designed for ASR, TTS, speaker diarization, and conversational speech research. Languages (13) Assamese, Bengali, Gujarati, Hindi, Kannada, Malayalam, Marathi, Nepali, Odia, Punjabi, Tamil, Telugu, Urdu. Content Conversations… See the full description on the dataset page: https://huggingface.co/datasets/eQOURSE/multilingual-speech.

sourceHugging Facecc-by-4.0updated 3mo agoView on Hugging Face
3likes154downloads
settings

This repository belongs to eQOURSE on Hugging Face.

Team Ai never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

namemultilingual-speech
visibilitypublic
licencecc-by-4.0
gatedno
ownereQOURSE
Account settings
eQOURSE/multilingual-speech · Team Ai