datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
tessera-quantization-research-evidence
Tessera Quantization Research Evidence
This dataset is the primary-source measurement evidence from an ongoing research
program studying calibrated low-bit quantization (ternary, int4, vector-quantized
codebooks) for LLM inference on heterogeneous AMD hardware (RDNA3 iGPU, XDNA1/2
NPU, Zen 4/5 CPU). The work is done in a fork of llama.cpp (project name
"Tessera") that adds calibrated per-tensor ternary/payload4/VQ quantization,
NPU offload, and RDNA3-native GPU kernels.
This is… See the full description on the dataset page: https://huggingface.co/datasets/Tribunus-dev/tessera-quantization-research-evidence.Algerian-STT-Cleaned-V5Algerian-STT-Super-Dataset-V2Algerian-STT-Cleaned-V3Musictaglog_asrAV-QuantBench-Dataset
AV-QuantBench
AV-QuantBench is a procedural audio-visual benchmark for evaluating multimodal foundation models on abstract temporal reasoning, cross-modal conflict detection, and synchronized data interpretation across finance, medical, and industrial domains.
This Hugging Face dataset repository is structured as a benchmark-style release. It contains:
split metadata in JSONL format,
question-answer annotations,
audio-visual sample assets,
manifest files by domain,
and… See the full description on the dataset page: https://huggingface.co/datasets/gfcfirefly/AV-QuantBench-Dataset.filtered_common_voice_tamil_english-preprocessed-quantizedfiltered_common_voice_tamil_english-preprocessed-quantized-v2ledger-briefsEng_clean_medium_ttsQuantumLifeIntelligenceMODULEQuantumGarden_56voice_clone_task
