datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
peft-blog-assetsAssets for PEFT blog posts
benchmark-graveyard
PEFT Benchmark Graveyard ⚰️
This is a place where various experiments, from various users, PEFT versions and experiment hardware are collected.
DO NOT ASSUME THAT THIS DATA IS CONSISTENT.
But it may be helpful in some ways that we don't know yet, so we collect it.
Structure
The JSON files in this dataset are files that come from the PEFT method comparison benchmark suite.
No other files are accepted.
There is minimal structure. Put the experiment results into the… See the full description on the dataset page: https://huggingface.co/datasets/peft-internal-testing/benchmark-graveyard.cat-image-datasetLabels (in this order):
sks cat sitting on a chair in front of a box of chocolates
sks cat playing on the Steam Deck
sks cat wearing a necklace while sitting in a box on a sofa
a box with four donuts in front of sks cat
sks cat wearing a pink veil with flowers on it and a dagger made out of yellow cardboard
sks cat between two pillows, with one pillow showing a polar bear and the other a fox
sks cat with an espresso reading the newspaper
a close up of a hand petting sks cat on the head
sks cat… See the full description on the dataset page: https://huggingface.co/datasets/peft-internal-testing/cat-image-dataset.OPD_PEFT-evalsoverlap-data-peftpeft-factoryeval_pi05-peft-so101-4tasks-aug_pick-place_40This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"robot_type": "so_follower",
"total_episodes": 20,
"total_frames": 39913,
"total_tasks": 1,
"chunks_size": 1000,
"data_files_size_in_mb": 100,
"video_files_size_in_mb": 200,
"fps": 30,
"splits": {
"train": "0:20"
},
"data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/hjkso1406/eval_pi05-peft-so101-4tasks-aug_pick-place_40.details_dfurman__llama-2-70b-dolphin-peft
Dataset Card for Evaluation run of dfurman/llama-2-70b-dolphin-peft
Dataset Summary
Dataset automatically created during the evaluation run of model dfurman/llama-2-70b-dolphin-peft on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_dfurman__llama-2-70b-dolphin-peft.router_PEFT_data_Math_self_generation_Qwen3-32Bbio-safety-peft-lora
CBRN Safety Alignment & PEFT-LoRA Fine-Tuning Dataset
This repository contains the synthetic instruction-tuning dataset (.jsonl) designed for parameter-efficient fine-tuning (PEFT-LoRA) of edge language models (specifically Qwen/Qwen2.5-1.5B-Instruct).
The dataset is curated to evaluate and modify model logit distributions, persona attributions, and dual-use safety boundaries regarding Chemical, Biological, Radiological, and Nuclear (CBRN) risk scenarios.
🤖 Dataset… See the full description on the dataset page: https://huggingface.co/datasets/devsgnr/bio-safety-peft-lora.CHIP2023-PromptCBLUE-pefteval_groot-peft-so101-4tasks-aug_pick-place_16This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"robot_type": "so_follower",
"total_episodes": 0,
"total_frames": 0,
"total_tasks": 0,
"chunks_size": 1000,
"data_files_size_in_mb": 100,
"video_files_size_in_mb": 200,
"fps": 30,
"splits": {},
"data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/hjkso1406/eval_groot-peft-so101-4tasks-aug_pick-place_16.details_dfurman__llama-2-13b-dolphin-peft
Dataset Card for Evaluation run of dfurman/llama-2-13b-dolphin-peft
Dataset Summary
Dataset automatically created during the evaluation run of model dfurman/llama-2-13b-dolphin-peft on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 5 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_dfurman__llama-2-13b-dolphin-peft.peft_merging_datapeft-unit-test-generation-experiments
PEFT Unit Test Generation Experiments
Dataset description
The PEFT Unit Test Generation Experiments dataset contains metadata and details about a set of trained models used for generating unit tests with parameter-efficient fine-tuning (PEFT) methods. This dataset includes models from multiple namespaces and various sizes, trained with different tuning methods to provide a comprehensive resource for unit test generation research.
Dataset Structure
Data… See the full description on the dataset page: https://huggingface.co/datasets/andstor/peft-unit-test-generation-experiments.details_jondurbin__airoboros-65b-gpt4-1.4-peft
Dataset Card for Evaluation run of jondurbin/airoboros-65b-gpt4-1.4-peft
Dataset Summary
Dataset automatically created during the evaluation run of model jondurbin/airoboros-65b-gpt4-1.4-peft on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_jondurbin__airoboros-65b-gpt4-1.4-peft.router_PEFT_data_Math_Qwen3-8Brouter_PEFT_data_Math_Qwen3-14Btsfm-peft-bench
TSFM-PEFT-Bench
A cross-architecture benchmark for evaluating Parameter-Efficient Fine-Tuning
(PEFT) recommendation reliability in Time Series Foundation Models (TSFMs).
Companion code and artifacts for the paper "TSFM-PEFT-Bench: A
Cross-Architecture Benchmark for PEFT Selection in Time Series Foundation
Models" (under double-blind review at NeurIPS 2026 Datasets and Benchmarks
Track).
Quick metadata:
License: Apache-2.0 (LICENSE)
Croissant manifest: tsfm_peft_bench.croissant.json… See the full description on the dataset page: https://huggingface.co/datasets/EvalData/tsfm-peft-bench.rollout_sort_b601_simple_filtered_smolvla_finetune_w_peft_v1This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"fps": 30,
"features": {
"action": {
"dtype": "float32",
"names": [
"shoulder_pan.pos",
"shoulder_lift.pos",
"elbow_flex.pos",
"wrist_flex.pos",
"wrist_yaw.pos",
"wrist_roll.pos",
"gripper.pos"
]… See the full description on the dataset page: https://huggingface.co/datasets/adrfm/rollout_sort_b601_simple_filtered_smolvla_finetune_w_peft_v1.details_dfurman__falcon-40b-openassistant-peft
Dataset Card for Evaluation run of dfurman/falcon-40b-openassistant-peft
Dataset Summary
Dataset automatically created during the evaluation run of model dfurman/falcon-40b-openassistant-peft on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_dfurman__falcon-40b-openassistant-peft.router_PEFT_data_Math_Qwen3-4Bpeft_test_safedetails_ferdinandjasong__SuperCoder-7B-Qwen2.5-0525-peft-merged
Dataset Card for Evaluation run of ferdinandjasong/SuperCoder-7B-Qwen2.5-0525-peft-merged
Dataset automatically created during the evaluation run of model ferdinandjasong/SuperCoder-7B-Qwen2.5-0525-peft-merged.
The dataset is composed of 2 configuration, each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/ferdinandjasong/details_ferdinandjasong__SuperCoder-7B-Qwen2.5-0525-peft-merged.ap-sql-peft
ap-sql-peft
A chat-format text-to-SQL dataset for Accounts Payable analytics on Oracle. The schema is modelled on Oracle Fusion
AP, Payments and Supplier tables, and the data behind it is synthetic. The dataset was used to train the LoRA adapter
samrat-kar/ap-sql-v1.
Every gold SQL query was run against the demo database when the dataset was built; meta.result_rows records how many rows it returned.
Format
One JSON object per line:
{"messages": [{"role": "system"… See the full description on the dataset page: https://huggingface.co/datasets/samrat-kar/ap-sql-peft.router_PEFT_data_Math_simple_prompt_Qwen3-32Bdetails_dfurman__llama-2-13b-guanaco-peft
Dataset Card for Evaluation run of dfurman/llama-2-13b-guanaco-peft
Dataset Summary
Dataset automatically created during the evaluation run of model dfurman/llama-2-13b-guanaco-peft on the Open LLM Leaderboard.
The dataset is composed of 61 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_dfurman__llama-2-13b-guanaco-peft.details_dfurman__Mixtral-8x7B-peft-v0.1
Dataset Card for Evaluation run of dfurman/Mixtral-8x7B-peft-v0.1
Dataset automatically created during the evaluation run of model dfurman/Mixtral-8x7B-peft-v0.1 on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_dfurman__Mixtral-8x7B-peft-v0.1.PEFT-Benchmarking-results
Reproducibility and Benchmarking of Parameter-Efficient Fine-Tuning Methods for Transformer Models
A reproducible empirical benchmark of Full Fine-Tuning, LoRA, AdaLoRA, Prefix Tuning, and IA³ across BERT-base and DistilBERT on three GLUE classification tasks.
While numerous PEFT methods have been proposed to reduce the cost of fine-tuning large transformer models, existing evaluations are often conducted under different experimental settings, making direct comparison… See the full description on the dataset page: https://huggingface.co/datasets/satyansh0/PEFT-Benchmarking-results.router_PEFT_data_Math_Qwen3-32B
