Team Ai
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01laion /openswe-tasks-patched-v5text10K<n<100K0 likes1.5k downloads5mo agoHugging Face02patched-codes /static-analysis-evalA dataset of 76 Python programs taken from real Python open source projects (top 100 on GitHub), where each program is a file that has exactly 1 vulnerability as detected by a particular static analyzer (Semgrep), used in the paper Patched MOA: optimizing inference for diverse software development tasks. OpenAI used the synth-vuln-fixes and fine-tuned a new version of gpt-4o is now the SOTA on this benchmark. More details and code is available from their repo. More details on the benchmark… See the full description on the dataset page: https://huggingface.co/datasets/patched-codes/static-analysis-eval.textn<1K20 likes879 downloads1y agoHugging Face03CIRCL /vulnerability-cwe-patch Description This dataset, CIRCL/vulnerability-cwe-patch, provides structured, real-world vulnerabilities enriched with CWE identifiers and corresponding patches from platforms like GitHub and GitLab. It is designed to support the development of tools for vulnerability classification, triage, and automated remediation. Each entry includes metadata such as CVE/GHSA ID, a description, CWE categorization, and links to verified patch commits with associated diff content and commit… See the full description on the dataset page: https://huggingface.co/datasets/CIRCL/vulnerability-cwe-patch.text1K<n<10K5 likes517 downloads3mo agoHugging Face04kajuma /diffllama_patch_tokenized1M<n<10M0 likes508 downloads8mo agoHugging Face05anaumghori /patchlet-embed-preprocessedimage100K<n<1M0 likes475 downloads8mo agoHugging Face06patched-codes /generate-readme-eval Generate README Eval The generate-readme-eval is a dataset (train split) and benchmark (test split) to evaluate the effectiveness of LLMs when summarizing entire GitHub repos in form of a README.md file. The datset is curated from top 400 real Python repositories from GitHub with at least 1000 stars and 100 forks. The script used to generate the dataset can be found here. For the dataset we restrict ourselves to GH repositories that are less than 100k tokens in size to allow us to… See the full description on the dataset page: https://huggingface.co/datasets/patched-codes/generate-readme-eval.textsummarizationn<1K3 likes417 downloads2y agoHugging Face07closji /cc12m_openai_clip-vit-base-patch32_image_image_retrieval_pairs_2022-09-13image10M<n<100M0 likes415 downloads4y agoHugging Face08rasdani /github-patches-genesys-swe-prompttext10K<n<100K0 likes332 downloads1y agoHugging Face09dpdl-benchmark /patch_camelyonimage100K<n<1M0 likes303 downloads2y agoHugging Face10kajuma /training_03_05_patch1M<n<10M0 likes269 downloads2y agoHugging Face11Martingkc /LLaVa-CC3M-Pretrain-clip-vit-base-patch32text100K<n<1M0 likes248 downloads6mo agoHugging Face12zacharielegault /PatchCamelyon PatchCamelyon (PCam) This is a reupload of the PatchCamelyon (PCam) dataset to make it more readily usable instead of manipulating H5 files. The original can be found in the author's Github repo. If you use this dataset, please cite the original publications: @inproceedings{veeling2018rotation, title={Rotation Equivariant CNNs for Digital Pathology}, author={Veeling, Bastiaan S and Linmans, Jasper and Winkens, Jim and Cohen, Taco and Welling, Max}, booktitle={Medical Image… See the full description on the dataset page: https://huggingface.co/datasets/zacharielegault/PatchCamelyon.imageimage-classification100K<n<1M2 likes246 downloads2y agoHugging Face13Martingkc /LLaVa-CC3M-PostTraining-clip-vit-base-patch16text100K<n<1M0 likes236 downloads6mo agoHugging Face14R2E-Gym /R2E-TestgenAgent-Patchestextn<1K1 likes226 downloads1y agoHugging Face15open-athena /a3-rl-DCAgent_r2egym-patched-full-oracletext10K<n<100K0 likes219 downloads4mo agoHugging Face16ai-sec-lab /PatchBench PatchBench PatchBench is a benchmark for evaluating AI agents on realistic vulnerability patching tasks: 213 tasks drawn from 32 popular GitHub C/C++ projects. It selects vulnerabilities whose ground-truth fixes lie outside the crash stack, and uses vulnerability transplant plus code mutation to mitigate surface-level fixes and patch memorization. This repository holds the task metadata, one row per task to identify the project, the exact repository state, the crash, and the… See the full description on the dataset page: https://huggingface.co/datasets/ai-sec-lab/PatchBench.texttext-generationn<1K0 likes182 downloads20d agoHugging Face17prakanda /SynthMat_Patches_DStext100K<n<1M0 likes176 downloads2y agoHugging Face18closji /cc12m_openai_clip-vit-base-patch32_image_image_retrieval_pairs_2022-09-15image10M<n<100M0 likes169 downloads4y agoHugging Face19DCAgent /swe_rebench_patchedtext1K<n<10K0 likes163 downloads7mo agoHugging Face20open-athena /rl__24GPU_shaped__swe_rebench_patched_oracle__r2egym-nl2bash-stacktext10K<n<100K0 likes159 downloads7mo agoHugging Face21michoo42 /Patchnoisseur Patchnoisseur A connoisseur's cellar of CVEs: every NVD CVE joined to its fixing-commit diff (when one could be found), its NVD description, and its associated CWE(s) (id, name, short description) — served as a single Parquet dataset. 351 884 CVEs · 25 015 with a real git diff attached · CVE-1999 → CVE-2026 · ~744 MB on disk (zstd-compressed Parquet, sharded ~300 MB each). What's in it One row per CVE in the NVD feed. CVEs without a retrievable patch are… See the full description on the dataset page: https://huggingface.co/datasets/michoo42/Patchnoisseur.tabulartext-classification100K<n<1M0 likes146 downloads5mo agoHugging Face22R2E-Gym /R2EGym-VerifierTrajectories-PatchOnlytext1K<n<10K0 likes139 downloads2y agoHugging Face23lighteternal /biodecision-sft-v2.2-patch A newer, better version is out: BioDecision v2 training data. BioDecision v2 is trained on 1.75M decisions from 57 sources (22 new) and scores 75.3% vs 69.4% for BioDecision v1 on 63,432 held-out biomedical test decisions. This BioDecision v1 page stays available. BioDecision SFT v2.2 patch 32,374 decisions for a short second training stage of BioDecision-4B. It targets two weaknesses measured after stage 1: judging whether an answer is supported by a source, and forecasting… See the full description on the dataset page: https://huggingface.co/datasets/lighteternal/biodecision-sft-v2.2-patch.texttext-classification10K<n<100K0 likes134 downloads5d agoHugging Face24DCAgent /rl__24GPU_base__swe_rebench_patched_oracle__r2egym-nl2bash-stacktext10K<n<100K0 likes132 downloads7mo agoHugging Face25Martingkc /LLaVa-Instruct-150K-clip-vit-base-patch32text100K<n<1M0 likes132 downloads6mo agoHugging Face261g0rrr /ny_test_grab_patched_wire_40This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v2.1", "robot_type": "sam_evt2", "total_episodes": 40, "total_frames": 22666, "total_tasks": 1, "total_videos": 160, "total_chunks": 1, "chunks_size": 1000, "fps": 60, "splits": { "train": "0:40" }, "data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/1g0rrr/ny_test_grab_patched_wire_40.tabularrobotics10K<n<100K0 likes125 downloads9mo agoHugging Face27rasdani /github-patchestext10K<n<100K0 likes122 downloads1y agoHugging Face28rasdani /github-patches-genesysimport re import json from datasets import load_dataset PROMPT_TEMPLATE = """\ We are currently solving the following issue within our repository. Here is the issue text: --- BEGIN ISSUE --- {issue} --- END ISSUE --- Below are some code segments, each from a relevant file. One or more of these files may contain bugs. --- BEGIN FILES --- {file_context} --- END FILES --- Please first localize the bug based on the issue statement, and then generate a patch according to the `git diff` format… See the full description on the dataset page: https://huggingface.co/datasets/rasdani/github-patches-genesys.text10K<n<100K0 likes122 downloads1y agoHugging Face29closji /cc12m_openai-clip-vit-patch32image1M<n<10M1 likes120 downloads4y agoHugging Face30open-athena /swe-rebench-patched-oracle-qwen3.5-122b-131k-opencode-literal-rescue-traces Agent trace dataset Decoding the literal token IDs The prompt_token_ids / completion_token_ids / logprobs columns are the verbatim tokens the serving engine emitted, stored PER AGENT STEP as a list-of-lists (one inner list per turn). To turn them back into text you MUST use the exact tokenizer the model was served with — a generic same-family tokenizer will decode word tokens to garbage. Served model / tokenizer source: Qwen/Qwen3.5-122B-A10B-FP8 from transformers… See the full description on the dataset page: https://huggingface.co/datasets/open-athena/swe-rebench-patched-oracle-qwen3.5-122b-131k-opencode-literal-rescue-traces.text1K<n<10K0 likes119 downloads2mo agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.