Team Ai
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01livecodebench /code_generation_liteLiveCodeBench is a temporaly updating benchmark for code generation. Please check the homepage: https://livecodebench.github.io/.n<1K111 likes92k downloads1y agoHugging Face02PRHW /loom-benchmark-livecodebench0 likes5.8k downloads3mo agoHugging Face03livecodebench /code_generation LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code 🏠 Home Page • 💻 GitHub Repository • 🏆 Leaderboard • LiveCodeBench is a "live" updating benchmark for holistically evaluating code related capabilities of LLMs. Particularly, it evaluates LLMs across a range of capabilties including code generation, self-repair, test output prediction, and code execution. This is the code generation scenario of LiveCodeBench. It is also… See the full description on the dataset page: https://huggingface.co/datasets/livecodebench/code_generation.textn<1K35 likes4.9k downloads2y agoHugging Face04QAQAQAQAQ /LiveCodeBench-Pro-Testcase0 likes4k downloads1y agoHugging Face05sam-paech /livecodebench-code_generation_litetext1K<n<10K0 likes2.4k downloads1y agoHugging Face06fjzzq2002 /impossible_livecodebenchtextn<1K1 likes1.9k downloads1y agoHugging Face07marianna13 /livecodebench_code_generation_lite_parquet LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code 🏠 Home Page • 💻 GitHub Repository • 🏆 Leaderboard • 📄 Paper Change Log Since LiveCodeBench is a continuously updated benchmark, we provide different versions of the dataset. Particularly, we provide the following versions of the dataset: release_v1: The initial release of the dataset with problems released between May 2023 and Mar 2024 containing 400… See the full description on the dataset page: https://huggingface.co/datasets/marianna13/livecodebench_code_generation_lite_parquet.text10K<n<100K0 likes1.6k downloads10mo agoHugging Face08QAQAQAQAQ /LiveCodeBench-Progatedtabular1K<n<10K10 likes1.1k downloads1y agoHugging Face09livecodebench /execution-v2tabularn<1K5 likes952 downloads2y agoHugging Face10namanbnsl /livecodebench-code_generation_litetext1K<n<10K5 likes757 downloads6mo agoHugging Face11livecodebench /test_generation LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code 🏠 Home Page • 💻 GitHub Repository • 🏆 Leaderboard • LiveCodeBench is a "live" updating benchmark for holistically evaluating code related capabilities of LLMs. Particularly, it evaluates LLMs across a range of capabilties including code generation, self-repair, test output prediction, and code execution. This is the code generation scenario of LiveCodeBench. It is also… See the full description on the dataset page: https://huggingface.co/datasets/livecodebench/test_generation.textn<1K8 likes677 downloads2y agoHugging Face12bzantium /livecodebench LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code Note: This is a clone of livecodebench/code_generation_lite updated to work with recent versions of the datasets library. The original repository uses a Python loading script which is no longer supported. This version provides the same data using the standard JSONL format for compatibility. Dataset Description LiveCodeBench is a "live" updating benchmark for holistically… See the full description on the dataset page: https://huggingface.co/datasets/bzantium/livecodebench.text1K<n<10K0 likes436 downloads10mo agoHugging Face13nvidia /LiveCodeBench-CPP LiveCodeBench-CPP: An Extension of LiveCodeBench for Contamination Free Evaluation in C++ Overview LiveCodeBench-CPP includes 454 problems from the release_v6 of LiveCodeBench, covering the period from October 2024 to May 2025. These problems are sourced from AtCoder (287 problems) and LeetCode (167 problems). AtCoder Problems: These require generated solutions to read inputs from standard input (stdin) and write outputs to standard output (stdout). For unit testing, the… See the full description on the dataset page: https://huggingface.co/datasets/nvidia/LiveCodeBench-CPP.textn<1K4 likes433 downloads1y agoHugging Face14livecodebench /execution Dataset Card for "livecodebench-execute" textn<1K6 likes324 downloads3y agoHugging Face15PrimeIntellect /LiveCodeBench-v5textn<1K0 likes291 downloads1y agoHugging Face16RewardGuided /livecodebench_code_generation_litetext1K<n<10K0 likes177 downloads4d agoHugging Face17drproduck /livecodebench-v6text1K<n<10K0 likes147 downloads1y agoHugging Face18BenchEvolver /livecodebench-plus LiveCodeBench-v6-Plus A curated coding benchmark of 91 problems selected by hardness/discrimination (lcb-v6-plus). It combines two sources, all in one clean schema: 64 evolved problems — mutated/evolved variants from LiveCodeBench-v6 (each carries its seed_problem). 27 original problems — un-evolved AtCoder problems taken directly from livecodebench/code_generation_lite release v6 (seed_problem is null). About BenchEvolver The evolved problems were produced by… See the full description on the dataset page: https://huggingface.co/datasets/BenchEvolver/livecodebench-plus.texttext-generationn<1K0 likes114 downloads4mo agoHugging Face19drproduck /qwen3-8b-livecodebench-subset-v6-n128textn<1K0 likes112 downloads1y agoHugging Face20livecodebench /submissions LiveCodeBench Submissions To submit your models to LiveCodeBench, you can now use this huggingface repository to directly upload model generations. To submit the model generations, simply drag and drop your model generations folder here and create a pull request. 1 likes108 downloads2y agoHugging Face21tokenintelligence /LiveCodeBench-SnapShot-0406 LiveCodeBench Official repository for the paper "LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code" 🏠 Home Page • 💻 Data • 🏆 Leaderboard • 🔍 Explorer Introduction LiveCodeBench provides holistic and contamination-free evaluation of coding capabilities of LLMs. Particularly, LiveCodeBench continuously collects new problems over time from contests across three competition platforms -- LeetCode, AtCoder… See the full description on the dataset page: https://huggingface.co/datasets/tokenintelligence/LiveCodeBench-SnapShot-0406.0 likes106 downloads6mo agoHugging Face22ali-elganzory /livecodebench-code_generation_litetext1K<n<10K0 likes100 downloads7mo agoHugging Face23Gen-Verse /LiveCodeBench0 likes95 downloads1y agoHugging Face24drproduck /qwen3-8b-livecodebench-v6-n128textn<1K0 likes83 downloads1y agoHugging Face25minimario /livecodebench-execute-v2text1K<n<10K1 likes71 downloads3y agoHugging Face26jakeatx /ATX-Swift-Qwen3.8-27B-Uncensored-IQ4_XS-M-LiveCodeBench-v6 LiveCodeBench v6 — ATX Swift Qwen3.8-27B Uncensored IQ4_XS-M Public reproducibility package for a four-seed direct code-generation evaluation of jakeatx/ATX-Swift-Qwen3.8-27B-Uncensored-IQ4_XS-M-GGUF. Result 357/400 = 89.25% pass@1 across four 100-task seeds. The arithmetic mean of the four seed rates is also 89.25%. Qwen's published BF16 LiveCodeBench v6 figure is 90.3%; this run is 1.05 percentage points lower. seed passed pass rate 0 90/100 90.00%… See the full description on the dataset page: https://huggingface.co/datasets/jakeatx/ATX-Swift-Qwen3.8-27B-Uncensored-IQ4_XS-M-LiveCodeBench-v6.text-generation0 likes71 downloads18d agoHugging Face27bzantium /ko-livecodebench Ko-LiveCodeBench This dataset is the livecodebench/code_generation_lite dataset with the question_content field translated to Korean. Dataset Versions The dataset provides multiple configurations (subsets) corresponding to different release versions: release_v1: Problems released between May 2023 and Mar 2024 (400 problems) release_v2: Problems released between May 2023 and May 2024 (511 problems) release_v3: Problems released between May 2023 and Jul 2024 (612 problems)… See the full description on the dataset page: https://huggingface.co/datasets/bzantium/ko-livecodebench.text1K<n<10K1 likes68 downloads10mo agoHugging Face28Sqwish /routerarena-livecodebench-tests0 likes68 downloads3mo agoHugging Face29nuprl /Ag-LiveCodeBench-XThis repository contains the multi-PL variant of LiveCodeBench, prepared in the Agnostics project. Find out more about the dataset and the related artifacts on the project website. The easiest way to benchmark a model on this dataset is with our scripts. textn<1K1 likes65 downloads1y agoHugging Face30Groq /LiveCodeBench-CodeGenerationtextquestion-answeringn<1K0 likes61 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.