Team Ai
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01quintic /llama_3.1-sae-23-29-code-activationstext10K<n<100K1 likes304 downloads2y agoHugging Face02akpsahan /Uncensored-CodeLlama Dataset Card for Evaluation run of ehartford/WizardLM-1.0-Uncensored-CodeLlama-34b Dataset Summary Dataset automatically created during the evaluation run of model ehartford/WizardLM-1.0-Uncensored-CodeLlama-34b on the Open LLM Leaderboard. The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named… See the full description on the dataset page: https://huggingface.co/datasets/akpsahan/Uncensored-CodeLlama.tabular100K<n<1M0 likes207 downloads5mo agoHugging Face03future-architect /Llama-3.3-Future-Code-Instructions Llama 3.3 Future Code Instructions Llama 3.3 Future Code Instructions is a large-scale instruction dataset synthesized with the Meta Llama 3.3 70B Instruct model. The dataset was generated with the method called Magpie, where we prompted the model to generate instructions likely to be asked by the users. In addition to the original prompt introduced by the authors, we conditioned the system prompt on what specific programming language the user has an interest in, gaining control… See the full description on the dataset page: https://huggingface.co/datasets/future-architect/Llama-3.3-Future-Code-Instructions.text1M<n<10M0 likes137 downloads1y agoHugging Face04PJMixers-Dev /nvidia_Llama-Nemotron-Post-Training-Dataset-v1-partial-codetext100K<n<1M0 likes120 downloads2y agoHugging Face05open-llm-leaderboard /EpistemeAI__Fireball-Meta-Llama-3.1-8B-Instruct-Agent-0.003-128K-code-ds-auto-detailsgated Dataset Card for Evaluation run of EpistemeAI/Fireball-Meta-Llama-3.1-8B-Instruct-Agent-0.003-128K-code-ds-auto Dataset automatically created during the evaluation run of model EpistemeAI/Fireball-Meta-Llama-3.1-8B-Instruct-Agent-0.003-128K-code-ds-auto The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 5 run(s). Each run can be found as a specific split in each configuration, the split being named… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI__Fireball-Meta-Llama-3.1-8B-Instruct-Agent-0.003-128K-code-ds-auto-details.tabular10K<n<100K4 likes62 downloads2y agoHugging Face06nyu-dice-lab /lm-eval-results-ajibawa-2023-Code-Llama-3-8B-private Dataset Card for Evaluation run of ajibawa-2023/Code-Llama-3-8B Dataset automatically created during the evaluation run of model ajibawa-2023/Code-Llama-3-8B The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 4 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-ajibawa-2023-Code-Llama-3-8B-private.tabular100K<n<1M0 likes60 downloads2y agoHugging Face07defog /wikisql_codellama_1000 Dataset Card for "wikisql_codellama_1000" More Information needed text1K<n<10K13 likes48 downloads3y agoHugging Face08ibranze /codellama_unity3d_v2text1K<n<10K6 likes46 downloads2y agoHugging Face09mlabonne /CodeLlama-2-20k CodeLlama-2-20k: A Llama 2 Version of CodeAlpaca This dataset is the sahil2801/CodeAlpaca-20k dataset with the Llama 2 prompt format described here. Here is the code I used to format it: from datasets import load_dataset # Load the dataset dataset = load_dataset('sahil2801/CodeAlpaca-20k') # Define a function to merge the three columns into one def merge_columns(example): if example['input']: merged = f"<s>[INST] <<SYS>>\nBelow is an instruction that describes a task… See the full description on the dataset page: https://huggingface.co/datasets/mlabonne/CodeLlama-2-20k.texttext-generation10K<n<100K16 likes43 downloads3y agoHugging Face10Asap7772 /code_contests_llamabase_mc_intermediate-part2-of-4tabular100K<n<1M0 likes41 downloads2y agoHugging Face11Asap7772 /code_contests_llamabase_mc_intermediatetabular1M<n<10M0 likes39 downloads2y agoHugging Face12open-llm-leaderboard /EpistemeAI2__Fireball-Meta-Llama-3.1-8B-Instruct-Agent-0.005-128K-code-COT-detailsgated Dataset Card for Evaluation run of EpistemeAI2/Fireball-Meta-Llama-3.1-8B-Instruct-Agent-0.005-128K-code-COT Dataset automatically created during the evaluation run of model EpistemeAI2/Fireball-Meta-Llama-3.1-8B-Instruct-Agent-0.005-128K-code-COT The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI2__Fireball-Meta-Llama-3.1-8B-Instruct-Agent-0.005-128K-code-COT-details.tabular10K<n<100K0 likes38 downloads2y agoHugging Face13sanjay920 /cortex-codellamatext1M<n<10M0 likes35 downloads3y agoHugging Face14luna-code /llamaindextext1K<n<10K0 likes35 downloads3y agoHugging Face15open-llm-leaderboard /EpistemeAI__Fireball-Meta-Llama-3.1-8B-Instruct-Agent-0.004-128K-code-COT-detailsgated Dataset Card for Evaluation run of EpistemeAI/Fireball-Meta-Llama-3.1-8B-Instruct-Agent-0.004-128K-code-COT Dataset automatically created during the evaluation run of model EpistemeAI/Fireball-Meta-Llama-3.1-8B-Instruct-Agent-0.004-128K-code-COT The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI__Fireball-Meta-Llama-3.1-8B-Instruct-Agent-0.004-128K-code-COT-details.tabular10K<n<100K0 likes35 downloads2y agoHugging Face16AIAT /Pangpuriye-generated_by_LLama3-codeLlama 🤖 Super AI Engineer Development Program Season 4 - Pangpuriye House - Generated by LLama3+codeLlama Pangpuriye's House Dataset - Generated Dataset from LLama3+codeLlama The dataset is a pack of text generation from LLama3 and codeLlama. The dataset is set under cc-by-nc-2.0 license. Content The dataset consists of 54,033 rows of input, instruction, and output. Most of the context in the dataset is in Thai. Whereas, the output is generally the answers regarding… See the full description on the dataset page: https://huggingface.co/datasets/AIAT/Pangpuriye-generated_by_LLama3-codeLlama.texttable-question-answering10K<n<100K0 likes32 downloads2y agoHugging Face17Asap7772 /code_contests_llamabase_mc_intermediate-part1-of-4tabular100K<n<1M0 likes32 downloads2y agoHugging Face18OliverYoung /codellama-threejstextn<1K3 likes31 downloads3y agoHugging Face19open-llm-leaderboard /EpistemeAI__Fireball-Meta-Llama-3.1-8B-Instruct-Agent-0.004-128K-code-ds-auto-detailsgated Dataset Card for Evaluation run of EpistemeAI/Fireball-Meta-Llama-3.1-8B-Instruct-Agent-0.004-128K-code-ds-auto Dataset automatically created during the evaluation run of model EpistemeAI/Fireball-Meta-Llama-3.1-8B-Instruct-Agent-0.004-128K-code-ds-auto The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI__Fireball-Meta-Llama-3.1-8B-Instruct-Agent-0.004-128K-code-ds-auto-details.tabular10K<n<100K0 likes30 downloads2y agoHugging Face20dhuynh95 /Magicoder-Evol-Instruct-500-CodeLlama-70b-tokenized-0.5-Special-Tokentextn<1K0 likes29 downloads3y agoHugging Face21MMEX /highway_code_llama3textn<1K0 likes29 downloads2y agoHugging Face22maveriq /spider-codellama-13b-instruct-hf-temp0.0-pandastext1K<n<10K0 likes29 downloads2y agoHugging Face23open-llm-leaderboard /EpistemeAI__Fireball-Meta-Llama-3.1-8B-Instruct-Agent-0.003-128K-code-detailsgated Dataset Card for Evaluation run of EpistemeAI/Fireball-Meta-Llama-3.1-8B-Instruct-Agent-0.003-128K-code Dataset automatically created during the evaluation run of model EpistemeAI/Fireball-Meta-Llama-3.1-8B-Instruct-Agent-0.003-128K-code The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI__Fireball-Meta-Llama-3.1-8B-Instruct-Agent-0.003-128K-code-details.tabular10K<n<100K0 likes29 downloads2y agoHugging Face24open-llm-leaderboard /EpistemeAI2__Fireball-Meta-Llama-3.1-8B-Instruct-Agent-0.003-128K-code-math-detailsgated Dataset Card for Evaluation run of EpistemeAI2/Fireball-Meta-Llama-3.1-8B-Instruct-Agent-0.003-128K-code-math Dataset automatically created during the evaluation run of model EpistemeAI2/Fireball-Meta-Llama-3.1-8B-Instruct-Agent-0.003-128K-code-math The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI2__Fireball-Meta-Llama-3.1-8B-Instruct-Agent-0.003-128K-code-math-details.tabular10K<n<100K0 likes29 downloads2y agoHugging Face25Asap7772 /code_contests_llamabase_mc_intermediate-part4-of-4tabular100K<n<1M0 likes29 downloads2y agoHugging Face26open-llm-leaderboard /EpistemeAI__Fireball-Meta-Llama-3.2-8B-Instruct-agent-003-128k-code-DPO-detailsgated Dataset Card for Evaluation run of EpistemeAI/Fireball-Meta-Llama-3.2-8B-Instruct-agent-003-128k-code-DPO Dataset automatically created during the evaluation run of model EpistemeAI/Fireball-Meta-Llama-3.2-8B-Instruct-agent-003-128k-code-DPO The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI__Fireball-Meta-Llama-3.2-8B-Instruct-agent-003-128k-code-DPO-details.tabular10K<n<100K0 likes28 downloads2y agoHugging Face27MIN12352 /codellama_java_pythontext1M<n<10M0 likes27 downloads2y agoHugging Face28open-llm-leaderboard /EpistemeAI__Polypsyche-Llama-3.1-8B-Instruct-Agent-0.003-128K-code-ds-auto-Empathy-detailsgated Dataset Card for Evaluation run of EpistemeAI/Polypsyche-Llama-3.1-8B-Instruct-Agent-0.003-128K-code-ds-auto-Empathy Dataset automatically created during the evaluation run of model EpistemeAI/Polypsyche-Llama-3.1-8B-Instruct-Agent-0.003-128K-code-ds-auto-Empathy The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI__Polypsyche-Llama-3.1-8B-Instruct-Agent-0.003-128K-code-ds-auto-Empathy-details.tabular10K<n<100K0 likes27 downloads2y agoHugging Face29SecCoderX /SecCoderX_CodeLlama_7b_GRPO_dataset Citation If you find our work helpful, feel free to give us a cite. @misc{wu2026securecodegenerationonline, title={Secure Code Generation via Online Reinforcement Learning with Vulnerability Reward Model}, author={Tianyi Wu and Mingzhe Du and Yue Liu and Chengran Yang and Terry Yue Zhuo and Jiaheng Zhang and See-Kiong Ng}, year={2026}, eprint={2602.07422}, archivePrefix={arXiv}, primaryClass={cs.CR}, url={https://arxiv.org/abs/2602.07422}… See the full description on the dataset page: https://huggingface.co/datasets/SecCoderX/SecCoderX_CodeLlama_7b_GRPO_dataset.text10K<n<100K0 likes26 downloads7mo agoHugging Face30maveriq /spider-codellama-70b-instruct-hf-temp0.0-pandastext1K<n<10K0 likes25 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.