Team Ai
7 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01RVtech /Audio2Tool Audio2Tool: Speak, Call, Act — A Dataset for Benchmarking Speech Tool Use Authors: Ramit Pahwa1,∗,∗∗, Apoorva Beedu1,∗, Parivesh Priye1, Rutu Gandhi†1, Saloni Takawale†1, Aruna Baijal1, Zengli Yang1 1 Rivian & Volkswagen Technologies &nbsp;·&nbsp; ∗ equal contribution &nbsp;·&nbsp; ∗∗ corresponding author &nbsp;·&nbsp; † equal contribution 📄 Project page / demo: https://audio2tool.github.io/ 📦 Dataset: https://huggingface.co/datasets/RVtech/Audio2Tool ✉️ Contact (corresponding… See the full description on the dataset page: https://huggingface.co/datasets/RVtech/Audio2Tool.audioautomatic-speech-recognition10K<n<100K3 likes3.3k downloads4mo agoHugging Face02nymtheescobar /bengali-talkshow-audio Bengali Talkshow Audio Dataset A large-scale collection of 1,180 Bengali talk show audio recordings totaling 789+ hours of multi-speaker speech, sourced from Bangladeshi television talk shows and political debate programs. Dataset Description This dataset contains audio from Bengali-language TV talk shows, political debates, and news discussion programs from major Bangladeshi television channels. Each recording features multiple speakers engaged in discussion, making it… See the full description on the dataset page: https://huggingface.co/datasets/nymtheescobar/bengali-talkshow-audio.audioaudio-classification1K<n<10K0 likes168 downloads8mo agoHugging Face03DatoricAI /audio-video-conversation-4000hgated Audio-Video Conversational Dataset 4,000 hours of synchronized speech and video of natural conversations: face movement, mouth motion, gestures, turn-taking, emotion, laughter, and interruptions across 20+ languages. This repository is a specification and preview listing. The production dataset is rights-cleared and delivered directly to buyers. Request access to see the full schema and get real samples. Overview The Audio-Video Conversational Dataset is a 4… See the full description on the dataset page: https://huggingface.co/datasets/DatoricAI/audio-video-conversation-4000h.textautomatic-speech-recognitionn<1K0 likes37 downloads3mo agoHugging Face04nyuuzyou /znanio-audios Dataset Card for Znanio.ru Educational Audio Dataset Summary This dataset contains 3,417 educational audio files from the znanio.ru platform, a resource for teachers, educators, students, and parents providing diverse educational content. Znanio.ru has been a pioneer in educational technologies and distance learning in the Russian-speaking internet since 2009. Languages The dataset is primarily in Russian, with potential multilingual content: Russian (ru): The… See the full description on the dataset page: https://huggingface.co/datasets/nyuuzyou/znanio-audios.textaudio-classification1K<n<10K0 likes25 downloads2y agoHugging Face05theblackcat102 /audio-alpacatexttext-generation10K<n<100K1 likes24 downloads3y agoHugging Face06OpenLLM-France /Luciole-Audio-Training-Dataset Luciole Audio Training Dataset Dataset description Luciole Audio Training Dataset is a large, multilingual, multi-task collection of audio–text conversations used to train the OpenLLM-France Luciole audio-language models. It adapts a text LLM to understand audio by pairing speech, music and environmental sounds with instruction-style dialogues (transcription, translation, spoken question answering, audio/music/sound captioning and question answering, speaker and… See the full description on the dataset page: https://huggingface.co/datasets/OpenLLM-France/Luciole-Audio-Training-Dataset.textautomatic-speech-recognition10M<n<100M1 likes6 downloads19d agoHugging Face07OpenLLM-France /Luciole-Audio-Evaluation-Dataset Luciole Audio Evaluation Dataset Dataset description Luciole Audio Evaluation Dataset is the held-out evaluation collection used to benchmark the OpenLLM-France Luciole audio-language models. It is the evaluation counterpart of the Luciole Audio Training Dataset: a multilingual, multi-task set of audio–text conversations, restricted to the test splits of each source dataset, covering transcription, speech translation, spoken and music question answering, music… See the full description on the dataset page: https://huggingface.co/datasets/OpenLLM-France/Luciole-Audio-Evaluation-Dataset.textautomatic-speech-recognition10K<n<100K1 likes6 downloads17d agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.