Team Ai
22 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01open-llm-leaderboard-old /details_google__codegemma-2b Dataset Card for Evaluation run of google/codegemma-2b Dataset automatically created during the evaluation run of model google/codegemma-2b on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_google__codegemma-2b.0 likes67 downloads2y agoHugging Face0211-47 /codegemma_gemini_pro_32_distilled_25k CodeGemma to Gemini Pro 3.2 Ultra-Advanced Code Distillation (25k) Dataset Description 25,000 unique, production-grade instruction-response pairs for distilling Gemini Pro 3.2-level coding intelligence into CodeGemma (or similar code models). Focus: Transfer frontier code reasoning — optimal algorithms, scalable system design, performance engineering, secure cryptography, and large-scale ML infrastructure — into smaller models. Dataset Structure… See the full description on the dataset page: https://huggingface.co/datasets/11-47/codegemma_gemini_pro_32_distilled_25k.text10K<n<100K3 likes55 downloads3mo agoHugging Face03open-llm-leaderboard /EpistemeAI2__Athene-codegemma-2-7b-it-alpaca-v1.2-detailsgated Dataset Card for Evaluation run of EpistemeAI2/Athene-codegemma-2-7b-it-alpaca-v1.2 Dataset automatically created during the evaluation run of model EpistemeAI2/Athene-codegemma-2-7b-it-alpaca-v1.2 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI2__Athene-codegemma-2-7b-it-alpaca-v1.2-details.tabular10K<n<100K0 likes51 downloads2y agoHugging Face04open-llm-leaderboard-old /details_google__codegemma-7b Dataset Card for Evaluation run of google/codegemma-7b Dataset automatically created during the evaluation run of model google/codegemma-7b on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_google__codegemma-7b.0 likes50 downloads2y agoHugging Face05open-llm-leaderboard /EpistemeAI__Athene-codegemma-2-7b-it-alpaca-v1.3-detailsgated Dataset Card for Evaluation run of EpistemeAI/Athene-codegemma-2-7b-it-alpaca-v1.3 Dataset automatically created during the evaluation run of model EpistemeAI/Athene-codegemma-2-7b-it-alpaca-v1.3 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI__Athene-codegemma-2-7b-it-alpaca-v1.3-details.tabular10K<n<100K0 likes48 downloads2y agoHugging Face06open-llm-leaderboard-old /details_google__codegemma-7b-it Dataset Card for Evaluation run of google/codegemma-7b-it Dataset automatically created during the evaluation run of model google/codegemma-7b-it on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_google__codegemma-7b-it.0 likes37 downloads2y agoHugging Face07open-llm-leaderboard /google__codegemma-1.1-2b-detailsgated0 likes32 downloads2y agoHugging Face081231czx /7b_codegemma_iter3text100K<n<1M0 likes30 downloads2y agoHugging Face09Ayush-Singh /reward-bench-codegemma-7b-it-yes-notabularn<1K0 likes25 downloads2y agoHugging Face10yyjb5 /tokenized_text_code_search_net_python_codegemma_bo1K<n<10K0 likes24 downloads2y agoHugging Face111231czx /code_gemma_raft_iter1text10K<n<100K0 likes13 downloads2y agoHugging Face12yyjb5 /mbpp-tokenized-codegemma-2b-0n<1K0 likes12 downloads2y agoHugging Face131231czx /7B_iter2_dpo_N1_random_pair_codegemmatext10K<n<100K0 likes12 downloads2y agoHugging Face141231czx /7b_codegemma_kto_iter2_random_pairtext10K<n<100K0 likes12 downloads2y agoHugging Face151231czx /7b_kto_iter2_mask_codegemma_2text100K<n<1M0 likes12 downloads2y agoHugging Face161231czx /kto_7b_iter2_mask_random_pair_codegemmatext10K<n<100K0 likes12 downloads2y agoHugging Face17Asap7772 /code_contests_codegemma_passk-part1-of-1textn<1K0 likes9 downloads2y agoHugging Face18yyjb5 /text_code_search_net_python_codegemma_botext100K<n<1M0 likes8 downloads2y agoHugging Face191231czx /7b_codegemma_kto_iter2text100K<n<1M0 likes8 downloads2y agoHugging Face20Asap7772 /code_contests_codegemma_passk-part1-of-1_gradedtextn<1K0 likes8 downloads2y agoHugging Face21yyjb5 /mbpp-tokenized-codegemma-2bn<1K0 likes4 downloads2y agoHugging Face221231czx /kto_7b_iter2_mask_codegemma_1text100K<n<1M0 likes4 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.