Team Ai
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01GeroldMeisinger /laion2b-en-a65_cogvlm2-4bit_captions Abstract This dataset contains image captions for the laion2B-en aesthetics>=6.5 image dataset using CogVLM2-4bit with the "laion-pop"-prompt to generate captions which were "likely" used in Stable Diffusion 3 training. From these image captions new synthetic images were generated using stable-diffusion-3-medium (batch-size=8). The synthetic images are best viewed locally by cloning this repo with: git lfs install git clone… See the full description on the dataset page: https://huggingface.co/datasets/GeroldMeisinger/laion2b-en-a65_cogvlm2-4bit_captions.imageimage-classification1K<n<10K6 likes3k downloads2y agoHugging Face02OALL /details_unsloth__llama-3-8b-bnb-4bit Dataset Card for Evaluation run of unsloth/llama-3-8b-bnb-4bit Dataset automatically created during the evaluation run of model unsloth/llama-3-8b-bnb-4bit. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_unsloth__llama-3-8b-bnb-4bit.tabular100K<n<1M0 likes134 downloads2y agoHugging Face03open-llm-leaderboard-old /details_Ramikan-BR__tinyllama_PY-CODER-4bit-lora_4k-v12 Dataset Card for Dataset Name This dataset card aims to be a base template for new datasets. It has been generated using this raw template. Dataset Details Dataset Description Curated by: [More Information Needed] Funded by [optional]: [More Information Needed] Shared by [optional]: [More Information Needed] Language(s) (NLP): [More Information Needed] License: [More Information Needed] Dataset Sources [optional] Repository: [More… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_Ramikan-BR__tinyllama_PY-CODER-4bit-lora_4k-v12.text-generation10K<n<100K0 likes133 downloads2y agoHugging Face04open-llm-leaderboard /sonthenguyen__ft-unsloth-zephyr-sft-bnb-4bit-20241014-170522-detailsgated Dataset Card for Evaluation run of sonthenguyen/ft-unsloth-zephyr-sft-bnb-4bit-20241014-170522 Dataset automatically created during the evaluation run of model sonthenguyen/ft-unsloth-zephyr-sft-bnb-4bit-20241014-170522 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sonthenguyen__ft-unsloth-zephyr-sft-bnb-4bit-20241014-170522-details.tabular10K<n<100K0 likes68 downloads2y agoHugging Face05open-llm-leaderboard /kms7530__chemeng_llama-3-8b-Instruct-bnb-4bit_24_1_100_1-detailsgated Dataset Card for Evaluation run of kms7530/chemeng_llama-3-8b-Instruct-bnb-4bit_24_1_100_1 Dataset automatically created during the evaluation run of model kms7530/chemeng_llama-3-8b-Instruct-bnb-4bit_24_1_100_1 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/kms7530__chemeng_llama-3-8b-Instruct-bnb-4bit_24_1_100_1-details.tabular10K<n<100K0 likes45 downloads2y agoHugging Face06open-llm-leaderboard-old /details_cloudyu__4bit_quant_TomGrc_FusionNet_34Bx2_MoE_v0.1_DPO Dataset Card for Evaluation run of cloudyu/4bit_quant_TomGrc_FusionNet_34Bx2_MoE_v0.1_DPO Dataset automatically created during the evaluation run of model cloudyu/4bit_quant_TomGrc_FusionNet_34Bx2_MoE_v0.1_DPO on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_cloudyu__4bit_quant_TomGrc_FusionNet_34Bx2_MoE_v0.1_DPO.0 likes33 downloads3y agoHugging Face07open-llm-leaderboard-old /details_robinsmits__Mistral-Instruct-7B-v0.2-ChatAlpacaV2-4bit Dataset Card for Evaluation run of robinsmits/Mistral-Instruct-7B-v0.2-ChatAlpacaV2-4bit Dataset automatically created during the evaluation run of model robinsmits/Mistral-Instruct-7B-v0.2-ChatAlpacaV2-4bit on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_robinsmits__Mistral-Instruct-7B-v0.2-ChatAlpacaV2-4bit.0 likes32 downloads3y agoHugging Face08open-llm-leaderboard-old /details_TFLai__llama-2-13b-4bit-alpaca-gpt4 Dataset Card for Evaluation run of TFLai/llama-2-13b-4bit-alpaca-gpt4 Dataset Summary Dataset automatically created during the evaluation run of model TFLai/llama-2-13b-4bit-alpaca-gpt4 on the Open LLM Leaderboard. The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_TFLai__llama-2-13b-4bit-alpaca-gpt4.0 likes30 downloads3y agoHugging Face09open-llm-leaderboard-old /details_Ramikan-BR__tinyllama-coder-py-4bit-v40 likes30 downloads2y agoHugging Face10open-llm-leaderboard /sonthenguyen__zephyr-sft-bnb-4bit-DPO-mtbc-213steps-detailsgated Dataset Card for Evaluation run of sonthenguyen/zephyr-sft-bnb-4bit-DPO-mtbc-213steps Dataset automatically created during the evaluation run of model sonthenguyen/zephyr-sft-bnb-4bit-DPO-mtbc-213steps The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sonthenguyen__zephyr-sft-bnb-4bit-DPO-mtbc-213steps-details.tabular10K<n<100K0 likes29 downloads2y agoHugging Face11open-llm-leaderboard-old /details_TFLai__pythia-2.8b-4bit-alpaca Dataset Card for Evaluation run of TFLai/pythia-2.8b-4bit-alpaca Dataset Summary Dataset automatically created during the evaluation run of model TFLai/pythia-2.8b-4bit-alpaca on the Open LLM Leaderboard. The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_TFLai__pythia-2.8b-4bit-alpaca.0 likes28 downloads3y agoHugging Face12open-llm-leaderboard /sonthenguyen__ft-unsloth-zephyr-sft-bnb-4bit-20241014-161415-detailsgated Dataset Card for Evaluation run of sonthenguyen/ft-unsloth-zephyr-sft-bnb-4bit-20241014-161415 Dataset automatically created during the evaluation run of model sonthenguyen/ft-unsloth-zephyr-sft-bnb-4bit-20241014-161415 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sonthenguyen__ft-unsloth-zephyr-sft-bnb-4bit-20241014-161415-details.tabular10K<n<100K0 likes26 downloads2y agoHugging Face13open-llm-leaderboard /kms7530__chemeng_phi-3-mini-4k-instruct-bnb-4bit_16_4_100_1_nonmath-detailsgated Dataset Card for Evaluation run of kms7530/chemeng_phi-3-mini-4k-instruct-bnb-4bit_16_4_100_1_nonmath Dataset automatically created during the evaluation run of model kms7530/chemeng_phi-3-mini-4k-instruct-bnb-4bit_16_4_100_1_nonmath The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/kms7530__chemeng_phi-3-mini-4k-instruct-bnb-4bit_16_4_100_1_nonmath-details.tabular10K<n<100K0 likes26 downloads2y agoHugging Face14open-llm-leaderboard /unsloth__phi-4-unsloth-bnb-4bit-detailsgated Dataset Card for Evaluation run of unsloth/phi-4-unsloth-bnb-4bit Dataset automatically created during the evaluation run of model unsloth/phi-4-unsloth-bnb-4bit The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/unsloth__phi-4-unsloth-bnb-4bit-details.tabular10K<n<100K0 likes26 downloads2y agoHugging Face15open-llm-leaderboard-old /details_TFLai__gpt-neo-1.3B-4bit-alpaca Dataset Card for Evaluation run of TFLai/gpt-neo-1.3B-4bit-alpaca Dataset Summary Dataset automatically created during the evaluation run of model TFLai/gpt-neo-1.3B-4bit-alpaca on the Open LLM Leaderboard. The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_TFLai__gpt-neo-1.3B-4bit-alpaca.0 likes24 downloads3y agoHugging Face16ezoujoh /eval_RFT_context_DeepSeek-R1-0528-Qwen3-8B-unsloth-bnb-4bit_telecomtextn<1K0 likes24 downloads1y agoHugging Face17open-llm-leaderboard-old /details_TFLai__gpt-neox-20b-4bit-alpaca Dataset Card for Evaluation run of TFLai/gpt-neox-20b-4bit-alpaca Dataset Summary Dataset automatically created during the evaluation run of model TFLai/gpt-neox-20b-4bit-alpaca on the Open LLM Leaderboard. The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_TFLai__gpt-neox-20b-4bit-alpaca.0 likes23 downloads3y agoHugging Face18open-llm-leaderboard /sonthenguyen__zephyr-sft-bnb-4bit-DPO-mtbo-180steps-detailsgated Dataset Card for Evaluation run of sonthenguyen/zephyr-sft-bnb-4bit-DPO-mtbo-180steps Dataset automatically created during the evaluation run of model sonthenguyen/zephyr-sft-bnb-4bit-DPO-mtbo-180steps The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sonthenguyen__zephyr-sft-bnb-4bit-DPO-mtbo-180steps-details.tabular10K<n<100K0 likes23 downloads2y agoHugging Face19open-llm-leaderboard /sonthenguyen__zephyr-sft-bnb-4bit-DPO-mtbr-180steps-detailsgated Dataset Card for Evaluation run of sonthenguyen/zephyr-sft-bnb-4bit-DPO-mtbr-180steps Dataset automatically created during the evaluation run of model sonthenguyen/zephyr-sft-bnb-4bit-DPO-mtbr-180steps The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sonthenguyen__zephyr-sft-bnb-4bit-DPO-mtbr-180steps-details.tabular10K<n<100K0 likes23 downloads2y agoHugging Face20open-llm-leaderboard /unsloth__phi-4-bnb-4bit-detailsgated Dataset Card for Evaluation run of unsloth/phi-4-bnb-4bit Dataset automatically created during the evaluation run of model unsloth/phi-4-bnb-4bit The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/unsloth__phi-4-bnb-4bit-details.tabular10K<n<100K0 likes23 downloads2y agoHugging Face21kth8 /Qwen3.5-27B-AWQ-4bit-GPQA-Diamond-benchmarkBenchmark of cyankiwi/Qwen3.5-27B-AWQ-4bit against fingertap/GPQA-Diamond dataset. Accuracy: 76.3% with Python tool. Metric Value Correct 151 Incorrect 46 Errors 1 Total samples 198 Python tool calls 225 Total completion tokens 659,879 Raw stats: { "accuracy": 0.763, "correct": 151, "incorrect": 46, "error": 1, "total": 198, "python_tool_calls": 225, "completion_tokens": 659879 } tabularn<1K0 likes23 downloads6mo agoHugging Face22open-llm-leaderboard-old /details_TFLai__llama-13b-4bit-alpaca Dataset Card for Evaluation run of TFLai/llama-13b-4bit-alpaca Dataset Summary Dataset automatically created during the evaluation run of model TFLai/llama-13b-4bit-alpaca on the Open LLM Leaderboard. The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_TFLai__llama-13b-4bit-alpaca.0 likes22 downloads3y agoHugging Face23open-llm-leaderboard-old /details_Ramikan-BR__tinyllama_PY-CODER-4bit-lora_4k-v50 likes22 downloads2y agoHugging Face24open-llm-leaderboard /sonthenguyen__ft-unsloth-zephyr-sft-bnb-4bit-20241014-164205-detailsgated Dataset Card for Evaluation run of sonthenguyen/ft-unsloth-zephyr-sft-bnb-4bit-20241014-164205 Dataset automatically created during the evaluation run of model sonthenguyen/ft-unsloth-zephyr-sft-bnb-4bit-20241014-164205 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sonthenguyen__ft-unsloth-zephyr-sft-bnb-4bit-20241014-164205-details.tabular10K<n<100K0 likes22 downloads2y agoHugging Face25math-extraction-comp /sonthenguyen__zephyr-sft-bnb-4bit-DPO-mtbr-180stepstabular1K<n<10K0 likes22 downloads2y agoHugging Face26ezoujoh /eval_SFT_context_DeepSeek-R1-0528-Qwen3-8B-unsloth-bnb-4bit_telecomtextn<1K0 likes22 downloads1y agoHugging Face27open-llm-leaderboard-old /details_Ramikan-BR__tinyllama-coder-py-4bit-v100 likes21 downloads2y agoHugging Face28open-llm-leaderboard /insightfactory__Llama-3.2-3B-Instruct-unsloth-bnb-4bitlora_model-detailsgated Dataset Card for Evaluation run of insightfactory/Llama-3.2-3B-Instruct-unsloth-bnb-4bitlora_model Dataset automatically created during the evaluation run of model insightfactory/Llama-3.2-3B-Instruct-unsloth-bnb-4bitlora_model The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/insightfactory__Llama-3.2-3B-Instruct-unsloth-bnb-4bitlora_model-details.tabular10K<n<100K0 likes21 downloads2y agoHugging Face29open-llm-leaderboard-old /details_Enno-Ai__vigogne2-enno-13b-sft-lora-4bit Dataset Card for Evaluation run of Enno-Ai/vigogne2-enno-13b-sft-lora-4bit Dataset Summary Dataset automatically created during the evaluation run of model Enno-Ai/vigogne2-enno-13b-sft-lora-4bit on the Open LLM Leaderboard. The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_Enno-Ai__vigogne2-enno-13b-sft-lora-4bit.0 likes20 downloads3y agoHugging Face30open-llm-leaderboard /dnhkng__RYS-Huge-bnb-4bit-detailsgated Dataset Card for Evaluation run of dnhkng/RYS-Huge-bnb-4bit Dataset automatically created during the evaluation run of model dnhkng/RYS-Huge-bnb-4bit The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/dnhkng__RYS-Huge-bnb-4bit-details.tabular10K<n<100K0 likes20 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.