datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
audioset-dasheng-0.6b-emb
AudioSet DaSheng-0.6B embeddings
Mean-pooled, float16 embeddings of
danjacobellis/audioset_opus_24kbps
from mispeech/dasheng-0.6B.
Columns
path: source clip path (string)
label: source AudioSet label indices (list of int64)
emb: 1,280-dimensional fixed-size list of float16
Audio is decoded from the source Opus bytes, mixed to mono, and resampled to
16 kHz. The embedding is the model's documented outputdim=None output:
sigmoid applied to the mean of the final… See the full description on the dataset page: https://huggingface.co/datasets/quinnlue/audioset-dasheng-0.6b-emb.dv-presidential-speechDhivehi Presidential Speech is a Dhivehi speech dataset created from data extracted and
processed by [Sofwath](https://github.com/Sofwath) as part of a collection of Dhivehi
datasets found [here](https://github.com/Sofwath/DhivehiDatasets).
The dataset contains around 2.5 hrs (1 GB) of speech collected from Maldives President's Office
consisting of 7 speeches given by President Yaameen Abdhul Gayyoom.
