stcoats/Lok_Sabha_test
Lok Sabha Spoken English Corpus — ParlaSpeech-compatible pilot This eight-hour pilot follows the Hugging Face structure used by ParlaSpeech-style speech corpora, with a compact schema tailored to Lok Sabha data. The default configuration contains one accepted aligned audio segment per row, with embedded 16 kHz audio, verbatim ASR, an explicitly separate edited UCR passage, word timings, speaker metadata, and source-order fields. Configurations default: 1,195… See the full description on the dataset page: https://huggingface.co/datasets/stcoats/Lok_Sabha_test.
016
No card is published for this repository, or it could not be fetched from Hugging Face right now.
