Team Ai
Datasetpublicgated

stcoats/Lok_Sabha_test

Lok Sabha Spoken English Corpus — ParlaSpeech-compatible pilot This eight-hour pilot follows the Hugging Face structure used by ParlaSpeech-style speech corpora, with a compact schema tailored to Lok Sabha data. The default configuration contains one accepted aligned audio segment per row, with embedded 16 kHz audio, verbatim ASR, an explicitly separate edited UCR passage, word timings, speaker metadata, and source-order fields. Configurations default: 1,195… See the full description on the dataset page: https://huggingface.co/datasets/stcoats/Lok_Sabha_test.

sourceHugging Facecc-by-nc-4.0updated 5d agoView on Hugging Face
0likes16downloads
Dataset Card

No card is published for this repository, or it could not be fetched from Hugging Face right now.