Team Ai
Datasetpublic

AudioCC-Lab/PICSAFEv1

Speech Quality Test Labels PICSAFEv1 is a multi-source annotated test dataset for evaluating speech quality assessment and audio data filtering methods. It contains 10,728 audio samples drawn from 14 source datasets, with annotations from a vocabulary of 33 tags. These tags describe recording provenance, speech styles, speaking rate and pitch, speaker attributes, noise, reverberation, distortion, and transcript errors. These annotations support benchmarking quality metrics and… See the full description on the dataset page: https://huggingface.co/datasets/AudioCC-Lab/PICSAFEv1.

sourceHugging Faceotherupdated 16d agoView on Hugging Face
0likes135downloads
9 commits on main
93dd4e916d ago

Introduce PICSAFEv1 purpose and applications

chuang.li
bbe9e9917d ago

Document binary tag policies and add metadata conversion script

chuang.li
887e6ad17d ago

Document tag to binary label conversion

chuang.li
78f03cc22d ago

Update dataset README

chuang.li
bf2181022d ago

Add metadata manifest and source download links

chuang.li
64c694b22d ago

Remove audio data and document source downloads

chuang.li
9a1f23723d ago

Fix dataset card data files config

chuang.li
5ad61e023d ago

Add PICSAFEv1 audio dataset

chuang.li
318dca023d ago

initial commit

jeremery123lc