datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Minecraft-GLB2Schem-RepairPairs-v1
unfundedResearcher/Minecraft-GLB2Schem-RepairPairs-v1
Paired (generated input, ground-truth target) Minecraft schematics for training a
model that turns an approximate voxelisation into a real build.
What a sample is
Each sample is three files inside a WebDataset TAR shard:
File
Meaning
<id>.input.schem
GENERATED. Produced by voxelising the source .glb. Approximate and noisy.
<id>.target.schem
GROUND TRUTH. The original schematic, copied byte-for-byte… See the full description on the dataset page: https://huggingface.co/datasets/unfundedResearcher/Minecraft-GLB2Schem-RepairPairs-v1.GlBBQ
GlBBQ: Galician Bias Benchmark for Question Answering
Dataset Description
GlBBQ is a Galician adaptation of the BBQ (Bias Benchmark for QA), a benchmark designed to measure social bias in multiple-choice question answering (QA) systems. More specifically, GlBBQ is derived from EsBBQ, the Spanish adaptation of BBQ.
The dataset follows the BBQ framework, where models are evaluated under:
Ambiguous contexts, where the correct answer is unknown and models should avoid… See the full description on the dataset page: https://huggingface.co/datasets/proxectonos/GlBBQ.
