datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
laion2b-en-a65_cogvlm2-4bit_captions
Abstract
This dataset contains image captions for the laion2B-en aesthetics>=6.5 image dataset using CogVLM2-4bit with the "laion-pop"-prompt to generate captions which were "likely" used in Stable Diffusion 3 training. From these image captions new synthetic images were generated using stable-diffusion-3-medium (batch-size=8).
The synthetic images are best viewed locally by cloning this repo with:
git lfs install
git clone… See the full description on the dataset page: https://huggingface.co/datasets/GeroldMeisinger/laion2b-en-a65_cogvlm2-4bit_captions.details_unsloth__llama-3-8b-bnb-4bit
Dataset Card for Evaluation run of unsloth/llama-3-8b-bnb-4bit
Dataset automatically created during the evaluation run of model unsloth/llama-3-8b-bnb-4bit.
The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_unsloth__llama-3-8b-bnb-4bit.details_Ramikan-BR__tinyllama_PY-CODER-4bit-lora_4k-v12
Dataset Card for Dataset Name
This dataset card aims to be a base template for new datasets. It has been generated using this raw template.
Dataset Details
Dataset Description
Curated by: [More Information Needed]
Funded by [optional]: [More Information Needed]
Shared by [optional]: [More Information Needed]
Language(s) (NLP): [More Information Needed]
License: [More Information Needed]
Dataset Sources [optional]
Repository: [More… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_Ramikan-BR__tinyllama_PY-CODER-4bit-lora_4k-v12.sonthenguyen__ft-unsloth-zephyr-sft-bnb-4bit-20241014-170522-details
Dataset Card for Evaluation run of sonthenguyen/ft-unsloth-zephyr-sft-bnb-4bit-20241014-170522
Dataset automatically created during the evaluation run of model sonthenguyen/ft-unsloth-zephyr-sft-bnb-4bit-20241014-170522
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sonthenguyen__ft-unsloth-zephyr-sft-bnb-4bit-20241014-170522-details.kms7530__chemeng_llama-3-8b-Instruct-bnb-4bit_24_1_100_1-details
Dataset Card for Evaluation run of kms7530/chemeng_llama-3-8b-Instruct-bnb-4bit_24_1_100_1
Dataset automatically created during the evaluation run of model kms7530/chemeng_llama-3-8b-Instruct-bnb-4bit_24_1_100_1
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/kms7530__chemeng_llama-3-8b-Instruct-bnb-4bit_24_1_100_1-details.details_cloudyu__4bit_quant_TomGrc_FusionNet_34Bx2_MoE_v0.1_DPO
Dataset Card for Evaluation run of cloudyu/4bit_quant_TomGrc_FusionNet_34Bx2_MoE_v0.1_DPO
Dataset automatically created during the evaluation run of model cloudyu/4bit_quant_TomGrc_FusionNet_34Bx2_MoE_v0.1_DPO on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_cloudyu__4bit_quant_TomGrc_FusionNet_34Bx2_MoE_v0.1_DPO.details_robinsmits__Mistral-Instruct-7B-v0.2-ChatAlpacaV2-4bit
Dataset Card for Evaluation run of robinsmits/Mistral-Instruct-7B-v0.2-ChatAlpacaV2-4bit
Dataset automatically created during the evaluation run of model robinsmits/Mistral-Instruct-7B-v0.2-ChatAlpacaV2-4bit on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_robinsmits__Mistral-Instruct-7B-v0.2-ChatAlpacaV2-4bit.details_TFLai__llama-2-13b-4bit-alpaca-gpt4
Dataset Card for Evaluation run of TFLai/llama-2-13b-4bit-alpaca-gpt4
Dataset Summary
Dataset automatically created during the evaluation run of model TFLai/llama-2-13b-4bit-alpaca-gpt4 on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_TFLai__llama-2-13b-4bit-alpaca-gpt4.details_Ramikan-BR__tinyllama-coder-py-4bit-v4sonthenguyen__zephyr-sft-bnb-4bit-DPO-mtbc-213steps-details
Dataset Card for Evaluation run of sonthenguyen/zephyr-sft-bnb-4bit-DPO-mtbc-213steps
Dataset automatically created during the evaluation run of model sonthenguyen/zephyr-sft-bnb-4bit-DPO-mtbc-213steps
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sonthenguyen__zephyr-sft-bnb-4bit-DPO-mtbc-213steps-details.details_TFLai__pythia-2.8b-4bit-alpaca
Dataset Card for Evaluation run of TFLai/pythia-2.8b-4bit-alpaca
Dataset Summary
Dataset automatically created during the evaluation run of model TFLai/pythia-2.8b-4bit-alpaca on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_TFLai__pythia-2.8b-4bit-alpaca.sonthenguyen__ft-unsloth-zephyr-sft-bnb-4bit-20241014-161415-details
Dataset Card for Evaluation run of sonthenguyen/ft-unsloth-zephyr-sft-bnb-4bit-20241014-161415
Dataset automatically created during the evaluation run of model sonthenguyen/ft-unsloth-zephyr-sft-bnb-4bit-20241014-161415
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sonthenguyen__ft-unsloth-zephyr-sft-bnb-4bit-20241014-161415-details.kms7530__chemeng_phi-3-mini-4k-instruct-bnb-4bit_16_4_100_1_nonmath-details
Dataset Card for Evaluation run of kms7530/chemeng_phi-3-mini-4k-instruct-bnb-4bit_16_4_100_1_nonmath
Dataset automatically created during the evaluation run of model kms7530/chemeng_phi-3-mini-4k-instruct-bnb-4bit_16_4_100_1_nonmath
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/kms7530__chemeng_phi-3-mini-4k-instruct-bnb-4bit_16_4_100_1_nonmath-details.unsloth__phi-4-unsloth-bnb-4bit-details
Dataset Card for Evaluation run of unsloth/phi-4-unsloth-bnb-4bit
Dataset automatically created during the evaluation run of model unsloth/phi-4-unsloth-bnb-4bit
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/unsloth__phi-4-unsloth-bnb-4bit-details.details_TFLai__gpt-neo-1.3B-4bit-alpaca
Dataset Card for Evaluation run of TFLai/gpt-neo-1.3B-4bit-alpaca
Dataset Summary
Dataset automatically created during the evaluation run of model TFLai/gpt-neo-1.3B-4bit-alpaca on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_TFLai__gpt-neo-1.3B-4bit-alpaca.eval_RFT_context_DeepSeek-R1-0528-Qwen3-8B-unsloth-bnb-4bit_telecomdetails_TFLai__gpt-neox-20b-4bit-alpaca
Dataset Card for Evaluation run of TFLai/gpt-neox-20b-4bit-alpaca
Dataset Summary
Dataset automatically created during the evaluation run of model TFLai/gpt-neox-20b-4bit-alpaca on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_TFLai__gpt-neox-20b-4bit-alpaca.sonthenguyen__zephyr-sft-bnb-4bit-DPO-mtbo-180steps-details
Dataset Card for Evaluation run of sonthenguyen/zephyr-sft-bnb-4bit-DPO-mtbo-180steps
Dataset automatically created during the evaluation run of model sonthenguyen/zephyr-sft-bnb-4bit-DPO-mtbo-180steps
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sonthenguyen__zephyr-sft-bnb-4bit-DPO-mtbo-180steps-details.sonthenguyen__zephyr-sft-bnb-4bit-DPO-mtbr-180steps-details
Dataset Card for Evaluation run of sonthenguyen/zephyr-sft-bnb-4bit-DPO-mtbr-180steps
Dataset automatically created during the evaluation run of model sonthenguyen/zephyr-sft-bnb-4bit-DPO-mtbr-180steps
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sonthenguyen__zephyr-sft-bnb-4bit-DPO-mtbr-180steps-details.unsloth__phi-4-bnb-4bit-details
Dataset Card for Evaluation run of unsloth/phi-4-bnb-4bit
Dataset automatically created during the evaluation run of model unsloth/phi-4-bnb-4bit
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/unsloth__phi-4-bnb-4bit-details.Qwen3.5-27B-AWQ-4bit-GPQA-Diamond-benchmarkBenchmark of cyankiwi/Qwen3.5-27B-AWQ-4bit against fingertap/GPQA-Diamond dataset.
Accuracy: 76.3% with Python tool.
Metric
Value
Correct
151
Incorrect
46
Errors
1
Total samples
198
Python tool calls
225
Total completion tokens
659,879
Raw stats:
{
"accuracy": 0.763,
"correct": 151,
"incorrect": 46,
"error": 1,
"total": 198,
"python_tool_calls": 225,
"completion_tokens": 659879
}
details_TFLai__llama-13b-4bit-alpaca
Dataset Card for Evaluation run of TFLai/llama-13b-4bit-alpaca
Dataset Summary
Dataset automatically created during the evaluation run of model TFLai/llama-13b-4bit-alpaca on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_TFLai__llama-13b-4bit-alpaca.details_Ramikan-BR__tinyllama_PY-CODER-4bit-lora_4k-v5sonthenguyen__ft-unsloth-zephyr-sft-bnb-4bit-20241014-164205-details
Dataset Card for Evaluation run of sonthenguyen/ft-unsloth-zephyr-sft-bnb-4bit-20241014-164205
Dataset automatically created during the evaluation run of model sonthenguyen/ft-unsloth-zephyr-sft-bnb-4bit-20241014-164205
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sonthenguyen__ft-unsloth-zephyr-sft-bnb-4bit-20241014-164205-details.sonthenguyen__zephyr-sft-bnb-4bit-DPO-mtbr-180stepseval_SFT_context_DeepSeek-R1-0528-Qwen3-8B-unsloth-bnb-4bit_telecomdetails_Ramikan-BR__tinyllama-coder-py-4bit-v10insightfactory__Llama-3.2-3B-Instruct-unsloth-bnb-4bitlora_model-details
Dataset Card for Evaluation run of insightfactory/Llama-3.2-3B-Instruct-unsloth-bnb-4bitlora_model
Dataset automatically created during the evaluation run of model insightfactory/Llama-3.2-3B-Instruct-unsloth-bnb-4bitlora_model
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/insightfactory__Llama-3.2-3B-Instruct-unsloth-bnb-4bitlora_model-details.details_Enno-Ai__vigogne2-enno-13b-sft-lora-4bit
Dataset Card for Evaluation run of Enno-Ai/vigogne2-enno-13b-sft-lora-4bit
Dataset Summary
Dataset automatically created during the evaluation run of model Enno-Ai/vigogne2-enno-13b-sft-lora-4bit on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_Enno-Ai__vigogne2-enno-13b-sft-lora-4bit.dnhkng__RYS-Huge-bnb-4bit-details
Dataset Card for Evaluation run of dnhkng/RYS-Huge-bnb-4bit
Dataset automatically created during the evaluation run of model dnhkng/RYS-Huge-bnb-4bit
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/dnhkng__RYS-Huge-bnb-4bit-details.
