datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
osworld_tasks_filesDataset_STTnew_dataset_sttAgri_STT_Benchmarking_DatasetThis is a domain-specific, multilingual agricultural speech dataset with a primary focus on Hindi, Telugu, and Odia, designed for speech-to-text and automatic speech recognition (ASR) tasks. It features human-annotated transcriptions and is intended for benchmarking ASR model performance in real-world agricultural scenarios.
This paper presents a comprehensive benchmark of 10 ASR models for agricultural advisory use across Hindi, Telugu, and Odia, using 10,934 real-world Farmer.Chat audio… See the full description on the dataset page: https://huggingface.co/datasets/bullseye-4/Agri_STT_Benchmarking_Dataset.sttstt-correctionSTTSnew_dataset_stt_audiostt-testresultsStt_200_data
