Team Ai
5 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01mlabonne /CodeLlama-2-20k CodeLlama-2-20k: A Llama 2 Version of CodeAlpaca This dataset is the sahil2801/CodeAlpaca-20k dataset with the Llama 2 prompt format described here. Here is the code I used to format it: from datasets import load_dataset # Load the dataset dataset = load_dataset('sahil2801/CodeAlpaca-20k') # Define a function to merge the three columns into one def merge_columns(example): if example['input']: merged = f"<s>[INST] <<SYS>>\nBelow is an instruction that describes a task… See the full description on the dataset page: https://huggingface.co/datasets/mlabonne/CodeLlama-2-20k.texttext-generation10K<n<100K16 likes43 downloads3y agoHugging Face02Yobitel /meta-llama-llama-3-1-70b-instruct__code-generation-humaneval-mini__019e3b935675 meta-llama/Llama-3.1-70B-Instruct on code.generation.humaneval-mini (NVIDIA H100 80GB HBM3) Back to leaderboard Headline metrics Metric Value Unit N Samples 5 N Ok 5 Ok Rate 1 Pass At 1 0.8 Pass At 1 P05 0.2 Pass At 1 P50 1 Pass At 1 P95 1 Timeout Rate 0 TTFT P50 28.0842 ms Total P50 Ms 4338.3245 Tokens Out Total 1653 Run configuration Model: meta-llama/Llama-3.1-70B-Instruct @ unknown00 Engine: vllm vunknown… See the full description on the dataset page: https://huggingface.co/datasets/Yobitel/meta-llama-llama-3-1-70b-instruct__code-generation-humaneval-mini__019e3b935675.text-generationn<1K0 likes28 downloads5mo agoHugging Face03Yobitel /meta-llama-llama-3-1-8b-instruct__code-generation-humaneval-mini__019e3b30b8df meta-llama/Llama-3.1-8B-Instruct on code.generation.humaneval-mini (NVIDIA H100 80GB HBM3) Back to leaderboard Headline metrics Metric Value Unit N Samples 5 N Ok 5 Ok Rate 1 Pass At 1 1 Pass At 1 P05 1 Pass At 1 P50 1 Pass At 1 P95 1 Timeout Rate 0 TTFT P50 15.7678 ms Total P50 Ms 1949.1334 Tokens Out Total 1589 Run configuration Model: meta-llama/Llama-3.1-8B-Instruct @ unknown00 Engine: vllm vunknown… See the full description on the dataset page: https://huggingface.co/datasets/Yobitel/meta-llama-llama-3-1-8b-instruct__code-generation-humaneval-mini__019e3b30b8df.text-generationn<1K0 likes16 downloads5mo agoHugging Face04Fu01978 /llama-q-a-code llama-q-a-code Description This dataset consists of synthetic coding-focused question-and-answer pairs designed for training or fine-tuning models to become better coding assistants. The dataset covers a broad range of programming topics, including Python, JavaScript, SQL, C++, and Git. Disclaimer Limitations To ensure efficient generation, responses were generated with a max_new_tokens limit of 512. Consequently, some complex coding answers or long… See the full description on the dataset page: https://huggingface.co/datasets/Fu01978/llama-q-a-code.texttext-generationn<1K0 likes13 downloads7mo agoHugging Face05ShahzebKhoso /local-code-arena-mbpp-codellama_7b Local Code Arena Telemetry: MBPP Benchmark on Code Llama 7B This repository hosts the raw evaluation metrics, execution telemetry logs, and structural syntax outputs captured from running the Mostly Basic Python Problems (MBPP) benchmark against Meta's Code Llama 7B model. This specific partition documents the baseline performance of early-generation specialized code engines, establishing a vital chronological anchor point to measure modern post-training alignment improvements.… See the full description on the dataset page: https://huggingface.co/datasets/ShahzebKhoso/local-code-arena-mbpp-codellama_7b.texttext-generationn<1K1 likes5 downloads5mo agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.