Team Ai
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01anggars /neural-mathrock Neural Math Rock Multimodal Emotion Dataset Dataset Description This corpus is a large-scale multimodal emotion classification dataset specifically developed for Music Information Retrieval (MIR) and emotional computational analysis within complex musical genres, predominantly Math Rock and Midwest Emo. The dataset consists of exactly 4,000 distinct full-length tracks structured directly from the validated metadata registry. The primary objective of this corpus is… See the full description on the dataset page: https://huggingface.co/datasets/anggars/neural-mathrock.audioaudio-classification1K<n<10K1 likes950 downloads4mo agoHugging Face02MathLLMs /VoiceAssistant-Eval 🔥 VoiceAssistant-Eval: Benchmarking AI Assistants across Listening, Speaking, and Viewing [🌐 Homepage] [🔮 Visualization] [💻 Github] [📖 Paper] [📊 Leaderboard ] [📊 Detailed Leaderboard ] [📊 Roleplay Leaderboard ] 🚀 Data Usage from datasets import load_dataset for split in ['listening_general', 'listening_music', 'listening_sound', 'listening_speech', 'speaking_assistant', 'speaking_emotion', 'speaking_instruction_following'… See the full description on the dataset page: https://huggingface.co/datasets/MathLLMs/VoiceAssistant-Eval.textquestion-answering10K<n<100K12 likes779 downloads1y agoHugging Face03MathematicianNLPer /MoulSot-Full MoulSot-Full Dataset Dataset Summary MoulSot-Full is a large-scale Moroccan Darija speech dataset containing in total 1,500 hours of speech audio. From this extensive corpus, a high-quality subset of approximately 80 hours has been carefully curated and transcribed. It was built entirely from publicly available YouTube content across 51 diverse channels (including vlogs, podcasts, interviews, and commentary) to capture real-world Moroccan Darija, including natural… See the full description on the dataset page: https://huggingface.co/datasets/MathematicianNLPer/MoulSot-Full.audio10K<n<100K0 likes309 downloads5mo agoHugging Face04ysdede /khanacademy-turkish-mathaudio10K<n<100K5 likes99 downloads2y agoHugging Face05mathildebindslev /MiniProjectMLThis dataset is designed for training an audio classification model that identifies the type of trash being thrown into a bucket. The model classifies sounds into the following categories: Metal, Glass, Plastic, Cardboard, and Noise (non-trash-related sounds). The dataset was recorded and organized as part of an Edge Impulse project to create a system that sorts trash based on sound. Link to Edge Impulse: https://studio.edgeimpulse.com/public/556872/live textaudio-classificationn<1K1 likes86 downloads2y agoHugging Face06Mathani-Ayat /qaloon-reciter-dataset Qaloon Reciter — continuously maintained data Initial release: 0.1.0; schema 1.0.0. A private team working corpus for Qālūn ʿan Nāfiʿ ASR. Current scope: al-Fātiḥah + Juzʾ ʿAmma (78–114), not the whole Quran. Three reciters have been approved by the project owner for this team release: Huthaify, Husary and Dokali. This approval is not a claim of a verified source copyright license or acoustic alignment review. Included and excluded material 569 recordings per… See the full description on the dataset page: https://huggingface.co/datasets/Mathani-Ayat/qaloon-reciter-dataset.audioautomatic-speech-recognition1K<n<10K0 likes66 downloads5d agoHugging Face07Lexia-Labs /french-math-asr-benchmarkaudio1K<n<10K1 likes36 downloads8mo agoHugging Face08matheushmart /MusicBoxaudion<1K0 likes27 downloads3y agoHugging Face09abdallah11119 /math-speech-datasetaudio1K<n<10K0 likes24 downloads11mo agoHugging Face10luvox-ai /lumi_math_13-7gatedaudion<1K0 likes23 downloads1y agoHugging Face11leungtianle /huawei-mathaudio100K<n<1M0 likes22 downloads1y agoHugging Face12AAAI2025 /MathSpeechaudio1K<n<10K4 likes21 downloads2y agoHugging Face13Mathani-Ayat /qaloon-reciter-experiments Experimental only — Waleed and Trabulsi Private research inventory uploaded at the project owner's request. Not part of the approved training dataset. Both sources retain authorization_status: unauthorized; an experimental upload is not source redistribution permission. Waleed: 569 individually labelled WAVs, durations read from WAV headers; Fātiḥah uses a separate basmalah track and fused ending. Source labels have not been independently checked. Some source identity entries… See the full description on the dataset page: https://huggingface.co/datasets/Mathani-Ayat/qaloon-reciter-experiments.audioautomatic-speech-recognition1K<n<10K0 likes17 downloads8d agoHugging Face14mrsndmn /MathSpeech_whisper_transcribed_normalizedaudio1K<n<10K0 likes16 downloads1y agoHugging Face15mrsndmn /MathSpeech_whisper_transcribedaudio1K<n<10K0 likes15 downloads1y agoHugging Face16mathsx0 /Hindi50_3audion<1K0 likes14 downloads1y agoHugging Face17AhmedZaky1 /math-audioaudio1K<n<10K0 likes14 downloads11mo agoHugging Face18NandinhoVinicius /matheusaleixoaudion<1K0 likes13 downloads3y agoHugging Face19abby1492 /mathbridge-audio Dataset Card for MathBridge_Audio Dataset Description This dataset is an expansion of the MathBridge dataset by Kyudan. The original dataset gives translations from English text to LaTeX formatting with context before and after. Our dataset is a subset of 1,000 rows from MathBridge with added audio files of speakers reading the full text. Curated by: Abigail Pitcairn Funded by: NSF and University of Southern Maine AIIR Lab Language: English (en) License: open-source… See the full description on the dataset page: https://huggingface.co/datasets/abby1492/mathbridge-audio.audio1K<n<10K1 likes13 downloads1y agoHugging Face20matheusfranco /spokensquad_perturbacoesaudio10K<n<100K0 likes13 downloads5mo agoHugging Face21Faeu /mathew1audion<1K0 likes12 downloads3y agoHugging Face22anonymous1unknown /MathSpeechaudio1K<n<10K0 likes12 downloads2y agoHugging Face23MatheusMarquesEiras /AudiosPortuguesaudion<1K0 likes12 downloads1y agoHugging Face24GabrielTOP /Matheusaudion<1K0 likes11 downloads3y agoHugging Face25mathiascdss /asdaudion<1K0 likes11 downloads3y agoHugging Face26zhouzc513622 /MathSpeechaudio1K<n<10K0 likes11 downloads10mo agoHugging Face27dcml0714 /speech_mathgatedaudio1K<n<10K2 likes10 downloads1y agoHugging Face28mathsx0 /Hindi50audion<1K0 likes8 downloads1y agoHugging Face29mathsx0 /Hindi50_2audion<1K0 likes8 downloads1y agoHugging Face30SilverPlasmid /MathAllThis is the testmini set of MathAll. audion<1K0 likes8 downloads6mo agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.