dfa
Datasets
All datasets matching “dfa”DF-arrowGenImage_webp: Just convert all images of GenImage to webp format, just save disk space
GenImage_test: only contains raw GenImage test samples
GenImage: raw GenImage with all data
DFADD
DFADD — Diffusion and Flow-Matching Based Audio Deepfake Dataset
Benchmark-ready packaging of the DFADD test (eval) split (DFADD: The
Diffusion and Flow-Matching Based Audio Deepfake Dataset, arXiv 2409.08731), a
VCTK-derived dataset targeting the newest generation of high-quality TTS
spoofing built on diffusion and flow-matching synthesizers.
Overview
Binary classification: bonafide (genuine VCTK recordings) vs. spoof
(text-to-speech generated from VCTK… See the full description on the dataset page: https://huggingface.co/datasets/SpeechAntiSpoofingBenchmarks/DFADD.DFADDDFADD_MLAAD_DiffSSD_VoxCeleb2cran-packages
CRAN packages dataset
R and Rmd source codes for CRAN packages.
The dataset has been constructed using the following steps:
Downloaded latest version from all packages on CRAN (see last updated). The source code has been downloaded from the GitHub mirror.
Identified the licenses from each package from their DESCRIPTION file, and classified each of them into some license_code. See the licenses.csv file.
Extract R and Rmd source files from all packages and joined with the package… See the full description on the dataset page: https://huggingface.co/datasets/dfalbel/cran-packages.Nsynth_Compressed
