datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
AppSecBench
AppSecBench Dataset Card
Dataset Summary
AppSecBench is an original benchmark of 406 vulnerable/secure code pairs spanning 12 programming
languages, 18 frameworks, 34 vulnerability classes, and 5 difficulty levels. Each record is a
self-contained evaluation case: a vulnerable snippet, its secure counterpart, an exploit sketch,
and the "ground truth" a detector/model is expected to produce (CWE, OWASP, severity, CVSS 3.1,
explainability, fix, and… See the full description on the dataset page: https://huggingface.co/datasets/ismailtasdelen/AppSecBench.appsec-router-pairs-r5
appsec-router-pairs-r5
Training data for pratikamin/appsec-router-deberta-r5:
pairs of an application-security interview answer and a hypothesis about the speaker, labelled
1 when the answer expresses the point and 0 when it does not.
Entirely synthetic. Answers were generated by openai/gpt-oss-120b (Apache 2.0) to 86 authored
interview questions and their 378 follow-ups from appsecinterview.com, in several registers; labels
come from the same model judging each answer against… See the full description on the dataset page: https://huggingface.co/datasets/pratikamin/appsec-router-pairs-r5.
