datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
model-bending-knowledge-base
Model Bending Knowledge Base
This dataset records what happens when you bend the inside of a diffusion model. Bending means multiplying, rotating,
adding noise to or otherwise changing the activations of a layer while the model generates.
Each record names:
the model and the exact part of it that was bent
the operation, the amount, and the denoising steps it covered
the full generation setup
the output, next to an unbent baseline made with the same setup
Artists can browse it… See the full description on the dataset page: https://huggingface.co/datasets/abuzreq/model-bending-knowledge-base.global-mmlu-rephrased
global_mmlu (rephrased for base-model evaluation)
Global MMLU knowledge-MCQA items rewritten from question format into completion/cloze format for base (non-instruction-tuned) language model evaluation.
Base (non-instruction-tuned) language models often can't follow question-style
prompts like "What is the capital of Turkey?" -- that phrasing is suited to
instruction-tuned models. Each item here has been rewritten into a natural
completion prefix (e.g. "The capital of Turkey is… See the full description on the dataset page: https://huggingface.co/datasets/base-model-evals/global-mmlu-rephrased.belebele-rephrased
belebele (rephrased for base-model evaluation)
Belebele reading-comprehension items rewritten from question format into completion/cloze format for base (non-instruction-tuned) language model evaluation.
Base (non-instruction-tuned) language models often can't follow question-style
prompts like "What is the capital of Turkey?" -- that phrasing is suited to
instruction-tuned models. Each item here has been rewritten into a natural
completion prefix (e.g. "The capital of Turkey is… See the full description on the dataset page: https://huggingface.co/datasets/base-model-evals/belebele-rephrased.dataset__countdown2arg__qwen2.5-1.5b-I__BoN__altered__convos__entropy__base_modelbase_model_sprint
Base Model Metadata Sprint
Description
Join us in improving the discoverability and understanding of models on the Hugging Face Hub by adding base_model metadata! This sprint aims to enhance the information available for models derived from, fine-tuned on, or quantized versions of existing base models.
🤗 Strong contributions will win prizes!! 🤗
Why It Matters
Adding base_model metadata helps users:
Easily find models derived from specific architectures… See the full description on the dataset page: https://huggingface.co/datasets/librarian-bots/base_model_sprint.train_data_imdb_from_base_modelterminal_bench_2_rl_rl_config_24GPU_base_yaml_model_path_Qwen3_8B_train_data_exad50f134kazakh_speech_dataset_ksdKazakh Speech Dataset cleaned, converted to parquet and with uppercase_transcription made with gpt4o_api.
Dataset info:
813 Speakers
with 500 samples for 4 speakers
with 250 samples for 809 speakers
Male/female
555 Hours
Guides
Load data 1
Replace the export HF_HOME with your HF_HOME path
from datasets import load_dataset
# export HF_HOME="/data/vladimir_albrekht/hf_cache"
ds = load_dataset("SRP-base-model-training/kazakh_speech_dataset_ksd") # split ='test' or… See the full description on the dataset page: https://huggingface.co/datasets/SRP-base-model-training/kazakh_speech_dataset_ksd.kazakh_speech_corpus_2
Kazakh_speech_dataset_2
This dataset contains Kazakh_speech_dataset_2 from ISSAI but in parquet format.
Dataset info
645,860 Utterances
1194 Hours in total
Sources in each split:
test : {'tv_news', 'crowdsourced', 'radio', 'talkshow', 'parliament', 'tts', 'podcasts'}
train : {'tv_news', 'crowdsourced', 'radio', 'talkshow', 'parliament', 'tts', 'podcasts'}
validation : {'tv_news', 'crowdsourced', 'radio', 'talkshow', 'parliament', 'tts','podcasts'}
Guides… See the full description on the dataset page: https://huggingface.co/datasets/SRP-base-model-training/kazakh_speech_corpus_2.hub_models_with_base_model_infoterminal_bench_2_rl_rl_config_24GPU_base_yaml_model_path_Qwen3_8B_train_data_ex4144df60terminal_bench_2_rl_rl_config_24GPU_base_yaml_model_path_Qwen3_8B_train_data_exb065ee39terminal_bench_2_rl_rl_config_24GPU_base_yaml_model_path_Qwen3_8B_train_data_exb28b6468swebench_verified_random_100_folders_rl_rl_config_24GPU_base_yaml_model_path_Qw411ef330swebench_verified_random_100_folders_rl_rl_config_24GPU_base_yaml_model_path_Qw41d4d58aswebench_verified_random_100_folders_rl_rl_config_24GPU_base_yaml_model_path_Qw9784788adev_set_v2_rl_rl_config_24GPU_base_yaml_model_path_Qwen3_8B_train_data_exp_rpt_856e9deeSelf-J-score-wo-ref-base-lla31-8b-inst-model-lla-31-8b-inst-thre-1MATH_OOD_Test_D1_Base_Model_Eval_COTself_evolving_iter-models-qwen3-4b-base_math_0116_2024-v0tetlock-binary-data-20260625_base_model_analyzedhub_models_with_base_model_info
Dataset Card for Hugging Face Hub Models with Base Model Metadata
Dataset Details
This dataset contains a subset of possible metadata for models hosted on the Hugging Face Hub.
All of these models contain base_model metadata i.e. information about the model used for fine-tuning.
This data can be used for creating network graphs showing links between models on the Hub.
Dataset Description
Curated by: [More Information Needed]
Funded by [optional]: [More… See the full description on the dataset page: https://huggingface.co/datasets/librarian-bots/hub_models_with_base_model_info.flair-base-model-detection
Flair Base Model Detection
For detailed instructions of dataset generation process, please refer to this GIST.
uned_super_rag_base_modelD-ExpTracker__FinEval_16k_fulleval_3args_basemodel-acronym_5o__v1
Experiment Tracker: FinEval_16k_fulleval_3args_basemodel-acronym_5o
Experiment Description: Evaluation experiment for task acronym_5o from FinEval_16k_fulleval_3args_basemodel
Start Time: 2025-10-27T02:02:12.652182
Tracker Dataset: TAUR-dev/D-ExpTracker__FinEval_16k_fulleval_3args_basemodel-acronym_5o__v1
Stages Completed
Total stages: 1
Models Created
Dataset Configurations
This tracker dataset contains the following configurations with… See the full description on the dataset page: https://huggingface.co/datasets/TAUR-dev/D-ExpTracker__FinEval_16k_fulleval_3args_basemodel-acronym_5o__v1.self_evolving_iter-models-qwen3-4b-base_math_0120_1951-v0D-EVAL__standard_eval_v3__GRPO_basemodel_rl_grpo-rl_8k_tok_eval-eval_rlD-ExpTracker__FinEval_16k_fulleval_3args_basemodel-longmult_2dig__v1
Experiment Tracker: FinEval_16k_fulleval_3args_basemodel-longmult_2dig
Experiment Description: Evaluation experiment for task longmult_2dig from FinEval_16k_fulleval_3args_basemodel
Start Time: 2025-10-27T00:50:28.324661
Tracker Dataset: TAUR-dev/D-ExpTracker__FinEval_16k_fulleval_3args_basemodel-longmult_2dig__v1
Stages Completed
Total stages: 1
Models Created
Dataset Configurations
This tracker dataset contains the following… See the full description on the dataset page: https://huggingface.co/datasets/TAUR-dev/D-ExpTracker__FinEval_16k_fulleval_3args_basemodel-longmult_2dig__v1.Self-J-score-w-ref-ref-lla31-70b-inst-base-lla31-8b-inst-model-lla-31-8b-inst-thre-1D-ExpTracker__FinEval_16k_fulleval_3args_basemodel-longmult_3dig__v1
Experiment Tracker: FinEval_16k_fulleval_3args_basemodel-longmult_3dig
Experiment Description: Evaluation experiment for task longmult_3dig from FinEval_16k_fulleval_3args_basemodel
Start Time: 2025-10-27T01:07:20.942754
Tracker Dataset: TAUR-dev/D-ExpTracker__FinEval_16k_fulleval_3args_basemodel-longmult_3dig__v1
Stages Completed
Total stages: 1
Models Created
Dataset Configurations
This tracker dataset contains the following… See the full description on the dataset page: https://huggingface.co/datasets/TAUR-dev/D-ExpTracker__FinEval_16k_fulleval_3args_basemodel-longmult_3dig__v1.
