datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
nixpkgs-security-patches
nixpkgs-security-patches
Training dataset for fine-tuning LLMs on nixpkgs security patch generation. Derived from real merged security PRs in NixOS/nixpkgs.
Dataset Details
588 training examples / 66 eval examples (654 total)
Format: Multi-turn tool-calling conversations in ChatML JSONL
Each example is a realistic agent session: the model reads the package file, finds the upstream fix, computes hashes via tools, and submits the fix for approval
Hashes and URLs… See the full description on the dataset page: https://huggingface.co/datasets/adastracomputing/nixpkgs-security-patches.patch_db
PatchDB: A Large-Scale Security Patch Dataset
Description
To foster large-scale research on vulnerability mitigation and to enable a comparison of different detection approaches, we make our dataset PatchDB from our DSN'21 paper publicly available.
PatchDB is a large-scale security patch dataset that contains around 12,073 security patches and 23,742 non-security patches from the real world.
You can find more details on the dataset in the paper "PatchDB: A Large-Scale… See the full description on the dataset page: https://huggingface.co/datasets/sunlab/patch_db.patchpilot-patchgen
PatchPilot patch-generation dataset
Supervised fine-tuning chats for PatchPilot's patch generator (A2). Each chat is exactly the
prompt PatchPilot's agent sends to its model, followed by the developers' real fix written in the
agent's SEARCH/REPLACE edit format.
Source
Built by scripts/build_patchgen_data.py (seed 42) from the SWE-bench training split
(princeton-nlp/SWE-bench, train) and the gold files' contents from princeton-nlp/SWE-bench_oracle.
The same seeded… See the full description on the dataset page: https://huggingface.co/datasets/Tejaswiniprabhakaran19/patchpilot-patchgen.ptdbench-llama-dapo-implementation-task-monkey-patch-011-dataset
PTDBench dataset snapshot: task_monkey_patch_011
This repository stores the immutable runtime dataset snapshot for one
materialized PTDBench task. It intentionally excludes model weights and
training checkpoints.
PTDBench family: llama_dapo_implementation
Source evaluation metric: val-core/math_dapo/acc/mean@1
Provenance: Processed from BytedTsinghua-SIA/DAPO-Math-17k; task-specific bytes are pinned.
License: Apache-2.0
The artifact manifest records every hydrated runtime path… See the full description on the dataset page: https://huggingface.co/datasets/LIF1014/ptdbench-llama-dapo-implementation-task-monkey-patch-011-dataset.
