datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Audio2Tool
Audio2Tool: Speak, Call, Act — A Dataset for Benchmarking Speech Tool Use
Authors: Ramit Pahwa1,∗,∗∗, Apoorva Beedu1,∗, Parivesh Priye1, Rutu Gandhi†1, Saloni Takawale†1, Aruna Baijal1, Zengli Yang1
1 Rivian & Volkswagen Technologies · ∗ equal contribution · ∗∗ corresponding author · † equal contribution
📄 Project page / demo: https://audio2tool.github.io/
📦 Dataset: https://huggingface.co/datasets/RVtech/Audio2Tool
✉️ Contact (corresponding… See the full description on the dataset page: https://huggingface.co/datasets/RVtech/Audio2Tool.bengali-talkshow-audio
Bengali Talkshow Audio Dataset
A large-scale collection of 1,180 Bengali talk show audio recordings totaling 789+ hours of multi-speaker speech, sourced from Bangladeshi television talk shows and political debate programs.
Dataset Description
This dataset contains audio from Bengali-language TV talk shows, political debates, and news discussion programs from major Bangladeshi television channels. Each recording features multiple speakers engaged in discussion, making it… See the full description on the dataset page: https://huggingface.co/datasets/nymtheescobar/bengali-talkshow-audio.audio-video-conversation-4000h
Audio-Video Conversational Dataset
4,000 hours of synchronized speech and video of natural conversations: face movement, mouth motion, gestures, turn-taking, emotion, laughter, and interruptions across 20+ languages.
This repository is a specification and preview listing. The production dataset is rights-cleared and delivered directly to buyers. Request access to see the full schema and get real samples.
Overview
The Audio-Video Conversational Dataset is a 4… See the full description on the dataset page: https://huggingface.co/datasets/DatoricAI/audio-video-conversation-4000h.znanio-audios
Dataset Card for Znanio.ru Educational Audio
Dataset Summary
This dataset contains 3,417 educational audio files from the znanio.ru platform, a resource for teachers, educators, students, and parents providing diverse educational content. Znanio.ru has been a pioneer in educational technologies and distance learning in the Russian-speaking internet since 2009.
Languages
The dataset is primarily in Russian, with potential multilingual content:
Russian (ru): The… See the full description on the dataset page: https://huggingface.co/datasets/nyuuzyou/znanio-audios.audio-alpacaLuciole-Audio-Training-Dataset
Luciole Audio Training Dataset
Dataset description
Luciole Audio Training Dataset is a large, multilingual, multi-task collection of
audio–text conversations used to train the OpenLLM-France
Luciole audio-language models. It adapts a text LLM to understand audio by pairing speech,
music and environmental sounds with instruction-style dialogues (transcription, translation,
spoken question answering, audio/music/sound captioning and question answering, speaker and… See the full description on the dataset page: https://huggingface.co/datasets/OpenLLM-France/Luciole-Audio-Training-Dataset.Luciole-Audio-Evaluation-Dataset
Luciole Audio Evaluation Dataset
Dataset description
Luciole Audio Evaluation Dataset is the held-out evaluation collection used to
benchmark the OpenLLM-France Luciole
audio-language models. It is the evaluation counterpart of the
Luciole Audio Training Dataset:
a multilingual, multi-task set of audio–text conversations, restricted to the test
splits of each source dataset, covering transcription, speech translation, spoken and
music question answering, music… See the full description on the dataset page: https://huggingface.co/datasets/OpenLLM-France/Luciole-Audio-Evaluation-Dataset.
