Team Ai
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01maxseats /aihub-464-preprocessed-680GB-set-52audio10K<n<100K0 likes711 downloads2y agoHugging Face02maxseats /aihub-464-preprocessed-680GB-set-56audio10K<n<100K0 likes503 downloads2y agoHugging Face03maxseats /aihub-464-preprocessed-680GB-set-57audio10K<n<100K0 likes448 downloads2y agoHugging Face04TheAIchemist13 /gramvaani_preprocessed_hi_train Dataset Card for "gramvaani_preprocessed_hi_train" More Information needed audio10K<n<100K0 likes392 downloads3y agoHugging Face05maxseats /aihub-464-preprocessed-680GB-set-53audio10K<n<100K0 likes392 downloads2y agoHugging Face06Jayem-11 /mozilla_commonvoice_hackathon_preprocessed_train_batch_3 Dataset Card for "mozilla_commonvoice_hackathon_preprocessed_train_batch_3" More Information needed audio10K<n<100K0 likes390 downloads3y agoHugging Face07mosama /sada-train-wav2vec2-xls-r-300m-ar-preprocessedaudio100K<n<1M0 likes380 downloads1y agoHugging Face08mosama /sada-train-preprocessedaudio100K<n<1M0 likes336 downloads1y agoHugging Face09Jayem-11 /mozilla_commonvoice_hackathon_preprocessed_train_batch_2 Dataset Card for "mozilla_commonvoice_hackathon_preprocessed_train_batch_2" More Information needed audio10K<n<100K0 likes310 downloads3y agoHugging Face10mteb /MUSIC-AVQA_cls-preprocessedaudio1K<n<10K0 likes295 downloads8mo agoHugging Face11maxseats /aihub-464-preprocessed-680GB-set-45audio10K<n<100K0 likes285 downloads2y agoHugging Face12qaz159qaz159 /preprocessed_speech_datasetsaudio100K<n<1M0 likes284 downloads2y agoHugging Face13htdung167 /vin100h-preprocessed-v2audio10K<n<100K2 likes248 downloads3y agoHugging Face14baohuynhbk14 /vin100h-preprocessed-16k-whisper-large-v3audio10K<n<100K0 likes196 downloads3y agoHugging Face15colinb83 /ltx25_preprocessed LTX-Video 2.5 Preprocessed Dataset Preprocessed training data for LTX-Video 2.5 (joint audio + video), stored as PyTorch tensors. Every file is a torch.save'd dict and can be loaded with: import torch d = torch.load("latents/10s/clip000001.pt", map_location="cpu", weights_only=True) Structure Clips are split by duration bucket (5s, 10s) and modality. File names are shared across folders: clipNNNNNN.pt refers to the same source clip in every folder that contains… See the full description on the dataset page: https://huggingface.co/datasets/colinb83/ltx25_preprocessed.video0 likes179 downloads15d agoHugging Face16EYEDOL /mozilla_commonvoice_naijaHausa1_preprocessed_train_batch_1audio10K<n<100K0 likes164 downloads1y agoHugging Face17EYEDOL /mozilla_commonvoice_Swahili_preprocessed_train_batch_1audio10K<n<100K0 likes164 downloads1y agoHugging Face18tankalapavankalyan /ds007808-sub01-speechopen-pangolin-preprocessed ds007808 sub-01 / speechopen / pangolin — preprocessed EEG↔speech windows Ready-to-train EEG↔speech windows for replicating the scaling experiment of Sato et al. 2024, "Scaling Law in Neural Data: Non-Invasive Speech Decoding with 175 Hours of EEG Data" (arXiv:2407.07595), built from the public ds007808 dataset (arXiv:2606.01264). Slice = subject sub-01, task speechopen (overt speech), device pangolin (128-ch g.Pangolin) — the rig matching the 175 h paper. Each example is one… See the full description on the dataset page: https://huggingface.co/datasets/tankalapavankalyan/ds007808-sub01-speechopen-pangolin-preprocessed.audioautomatic-speech-recognition100K<n<1M0 likes154 downloads4mo agoHugging Face19EYEDOL /mozilla_commonvoice_Swahili_preprocessed_train_batch_4audio10K<n<100K0 likes148 downloads1y agoHugging Face20EYEDOL /mozilla_commonvoice_Swahili_preprocessed_train_batch_3audio10K<n<100K0 likes146 downloads1y agoHugging Face21Jayem-11 /mozilla_commonvoice_hackathon_preprocessed_train_batch_6 Dataset Card for "mozilla_commonvoice_hackathon_preprocessed_train_batch_6" More Information needed audio10K<n<100K0 likes136 downloads3y agoHugging Face22EYEDOL /mozilla_commonvoice_Swahili_preprocessed_train_batch_2audio10K<n<100K0 likes136 downloads1y agoHugging Face23Jayem-11 /mozilla_commonvoice_hackathon_preprocessed_train_batch_4 Dataset Card for "mozilla_commonvoice_hackathon_preprocessed_train_batch_4" More Information needed audio10K<n<100K0 likes128 downloads3y agoHugging Face24Jayem-11 /mozilla_commonvoice_hackathon_preprocessed_train_batch_1 Dataset Card for "mozilla_commonvoice_hackathon_preprocessed_train_batch_1" More Information needed audio10K<n<100K0 likes119 downloads3y agoHugging Face25DewiBrynJones /preprocessed-whisper-btb-cv-cvad-wlga-ca-2607 Dataset Card Preprocessed Dataset: DewiBrynJones/preprocessed-whisper-btb-cv-cvad-wlga-ca-2607 Revision: main Dataset Statistics Train Split Statistics Dataset Revision Split Duration (HH:MM:SS) Clips Words Words/Clip % DewiBrynJones/banc-trawsgrifiadau-bangor-2605 5bfe2d098c8486d97fac8be76d86ec9146435245 train 56:46:32 50,557 589,095 11.7 31.9 techiaith/corpws-clllc-wlga 5d00294c31c78b1d7937bb2c2bc6cc70bc18d410 clips 48:20:49 27,579… See the full description on the dataset page: https://huggingface.co/datasets/DewiBrynJones/preprocessed-whisper-btb-cv-cvad-wlga-ca-2607.tabularautomatic-speech-recognition100K<n<1M0 likes112 downloads2mo agoHugging Face26Jayem-11 /mozilla_commonvoice_hackathon_preprocessed_train_batch_5 Dataset Card for "mozilla_commonvoice_hackathon_preprocessed_train_batch_5" More Information needed audio10K<n<100K0 likes110 downloads3y agoHugging Face27maxseats /aihub-464-preprocessed-680GB-set-0 필요 없는 컬럼 제거가 필요해요. 다음의 코드를 통해 만들었어요. import os import json from pydub import AudioSegment from tqdm import tqdm import re from datasets import Audio, Dataset, DatasetDict, load_from_disk, concatenate_datasets from transformers import WhisperFeatureExtractor, WhisperTokenizer import pandas as pd # 사용자 지정 변수를 설정해요. # DATA_DIR = '/mnt/a/maxseats/(주의-원본-680GB)주요 영역별 회의 음성인식 데이터' # 데이터셋이 저장된 폴더 DATA_DIR = '/mnt/a/maxseats/(주의-원본)split_files/set_0' # 첫 10GB 테스트 # 원천, 라벨링 데이터 폴더 지정… See the full description on the dataset page: https://huggingface.co/datasets/maxseats/aihub-464-preprocessed-680GB-set-0.audio10K<n<100K0 likes108 downloads2y agoHugging Face28Vano04 /MELD-Preprocessed MELD Preprocessed for SER This dataset is the manually preprocessed audio only version of MELD, only audio IDs, utterance transcriptions, dialogue IDs and Utterance IDs were extracted. S. Poria, D. Hazarika, N. Majumder, G. Naik, R. Mihalcea, E. Cambria. MELD: A Multimodal Multi-Party Dataset for Emotion Recognition in Conversation. (2018) Chen, S.Y., Hsu, C.C., Kuo, C.C. and Ku, L.W. EmotionLines: An Emotion Corpus of Multi-Party Conversations. arXiv preprint arXiv:1802.08379… See the full description on the dataset page: https://huggingface.co/datasets/Vano04/MELD-Preprocessed.audio10K<n<100K0 likes99 downloads10mo agoHugging Face29EYEDOL /mozilla_commonvoice_Swahili_preprocessed_train_batch_6audio1K<n<10K0 likes94 downloads1y agoHugging Face30mosama /sada-validation-preprocessed Details This is the SADA 2022 dataset with the input_features whish are log mels and the cleaned_labels which is the tokenized version of the cleaned_text. You can directly use this as the validation dataset when training Whisper Tiny, Small, Base & Medium models, as they all use the same tokenizer. Please double check this as well from the original model repo. In addtition, the following filters were applied to this data: All audios are less than 30 seconds and greater than 0… See the full description on the dataset page: https://huggingface.co/datasets/mosama/sada-validation-preprocessed.audioautomatic-speech-recognition1K<n<10K0 likes91 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.