Team Ai
24 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01evalstate /transformers-merge-experimentstabularn<1K3 likes1.2k downloads6mo agoHugging Face02baobabtech /evalexplorer-classify-experiments EvalExplorer document classifier: experiments The question When an evaluation report enters EvalExplorer, the ingestion pipeline sends its first pages to a large LLM (gpt-oss-120b, with Gemini 2.5 Flash and Qwen 3 235B as fallbacks), which returns five labels: evaluation approach (mixed methods, experimental, ...), type (impact evaluation, systematic review, ...), timing (baseline, midterm, endline), themes (global health, governance, ...) and countries (ISO… See the full description on the dataset page: https://huggingface.co/datasets/baobabtech/evalexplorer-classify-experiments.tabular1K<n<10K0 likes991 downloads4d agoHugging Face03tigerking009 /zoya-image-1-experiments ZOYA IMAGE-1 — Reproducible GGUF Experiments Purpose This dataset stores reproducible ZOYA IMAGE-1 image-generation experiments together with the exact generation parameters, model identities, SHA256 fingerprints, and validation reports. The package is designed for controlled comparisons where the tested variable is changed explicitly and all other relevant variables remain fixed. Current baseline Experiment ID: ZOYA_PHASE0_BASELINE_00001… See the full description on the dataset page: https://huggingface.co/datasets/tigerking009/zoya-image-1-experiments.imageimage-to-imagen<1K0 likes164 downloads2mo agoHugging Face04Praful932 /abmelt-experiments-exp_20260220_130124textn<1K0 likes99 downloads8mo agoHugging Face05lguerdan /indeterminacy-experimentstextn<1K0 likes96 downloads1y agoHugging Face06EdyVision /praxa-behavioral-experiments Behavioral KV Experiments Multi-domain protocol and evaluation results for behavioral policy alignment on COMPASS scenarios. Each HF config is one domain bundle: domains/<industry>/<scenario>/ with frozen split IDs (no query text) plus optional results/<model_key>/<condition>/ run artifacts. Associated paper Rosado, E. J. (2026). RFDT: Representation-first decision training for behavioral policy adaptation [Preprint]. URL pending. Acknowledgments and… See the full description on the dataset page: https://huggingface.co/datasets/EdyVision/praxa-behavioral-experiments.text1K<n<10K0 likes57 downloads5h agoHugging Face07Praful932 /abmelt-experiments-exp_20260218_171120textn<1K0 likes54 downloads8mo agoHugging Face08Praful932 /abmelt-experiments-exp_20260220_125440textn<1K0 likes50 downloads8mo agoHugging Face09mncai /fm-model-experiments-data FM Model Experiments — Synthetic VLM Training Data (KO + EN) Annotation data produced while building a native-resolution Korean+English VLM (GLM-4.6V vision tower transplanted onto a frozen GLM-5.2 743B MoE decoder). Code + technical report: https://github.com/genonai/fm-model-experiments This repo contains ANNOTATIONS ONLY (.jsonl). No images are redistributed. Every row references an image by a relative path (data/...); obtain the images from the original sources listed below… See the full description on the dataset page: https://huggingface.co/datasets/mncai/fm-model-experiments-data.textvisual-question-answering1K<n<10K0 likes50 downloads3mo agoHugging Face10Praful932 /abmelt-experiments-exp_20260219_182250textn<1K0 likes47 downloads8mo agoHugging Face11rmems /grok-1-ternary-quant-experiments Grok-1 SAAQ quantization / route-preservation experiments Dataset author: Raul Montoya Cardenas (rmems) SAAQ stands for Spiking Adaptive Activity Quantization, a term coined by the dataset author. Attribution: Grok Build: Grok 4.5 (high) packaged the original 2026-08-10 dataset. Codex: GPT-5.6-Sol (OpenAI) implemented, executed, validated, and published the canonical issue #85 v4 evidence added on 2026-08-24. Personal research measuring route preservation when packing open… See the full description on the dataset page: https://huggingface.co/datasets/rmems/grok-1-ternary-quant-experiments.tabularothern<1K0 likes43 downloads2mo agoHugging Face12Praful932 /abmelt-experiments-exp_20260219_182408textn<1K0 likes39 downloads8mo agoHugging Face13Praful932 /abmelt-experiments-exp_20260216_181056textn<1K0 likes27 downloads8mo agoHugging Face14willchow66 /mmmlu-bias-experiments MMMLU Bias Experiments Dataset Dataset Description This dataset contains 12 carefully designed experiments to measure language bias and position bias in Large Language Models (LLMs) using multilingual pairwise judgments. Key Features 12 Experiments: 8 original + 4 position-swapped experiments 11,478 samples per experiment (137,736 total test cases) Deterministic wrong answers: Uses fixed rule wrong_index = (correct_index + 1) % 4 Perfect correspondence: Wrong… See the full description on the dataset page: https://huggingface.co/datasets/willchow66/mmmlu-bias-experiments.textquestion-answering100K<n<1M0 likes24 downloads11mo agoHugging Face15Praful932 /abmelt-experiments-exp_20260215_104855textn<1K0 likes21 downloads8mo agoHugging Face16Praful932 /abmelt-experiments-exp_20260216_181641textn<1K0 likes21 downloads8mo agoHugging Face17Praful932 /abmelt-experiments-exp_20260215_070527textn<1K0 likes20 downloads8mo agoHugging Face18Praful932 /abmelt-experiments-exp_20260215_062019textn<1K0 likes15 downloads8mo agoHugging Face19Praful932 /abmelt-experiments-exp_20260215_051916textn<1K0 likes12 downloads8mo agoHugging Face20Praful932 /abmelt-experiments-exp_20260215_052622textn<1K0 likes9 downloads8mo agoHugging Face21rausch /prepared_context_4_experiments_old_train_bert-base-uncasedtextn<1K0 likes6 downloads2y agoHugging Face22likely /experimentstext1M<n<10M0 likes5 downloads2y agoHugging Face23rausch /prepared_data_file_zero_shot_prompting_100Q_4_experiments.jsontextn<1K0 likes4 downloads2y agoHugging Face24imran-siddique /iatp-experimentstextn<1K0 likes3 downloads9mo agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.