Team Ai
19 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01jedisct1 /security-auditsA collection of agent traces generated with Swival (not Claude Code, despite what the HF interface currently shows), an agent designed for open-source models. These traces focus on security audits of opensource software. Sharing traces with Swival Swival can export full conversation traces with --trace-dir, which writes one <session_id>.jsonl file per session: swival "Fix the login bug" --trace-dir traces/ Those JSONL files use Swival's Claude Code compatible trace export, and… See the full description on the dataset page: https://huggingface.co/datasets/jedisct1/security-audits.tabulartext-generation10K<n<100K18 likes17k downloads4mo agoHugging Face02secmlr /llm-fv-security-targets LLM-FV Security Targets This dataset contains 889 independently validated, containerized security-agent targets produced by the ucsb-mlsec/llm-fv pipelines. Contents GitHub Global Security Advisories: 446 targets OSS-Fuzz: 442 targets PoC task support: 889 targets Exploit task support: 447 targets Patch task support: 447 targets Compressed bundle size: 42.41 GiB Vulnerability classes: {'logic_bug': 450, 'memory_vulnerability': 439} Primary languages: {'C': 118… See the full description on the dataset page: https://huggingface.co/datasets/secmlr/llm-fv-security-targets.tabulartext-generationn<1K0 likes1.1k downloads2d agoHugging Face03OpenClaw /clawhub-security-signals ClawHub Security Signals 🦀 ClawHub | 📝 OpenClaw Blog | 🤗 Hugging Face Blog | 📄 Paper | 📄 Pre-Print ClawHub Security Signals is a sanitized, MIT-licensed security-signals dataset for public OpenClaw agent skills. It captures how an agent-skill registry evaluates trust, provenance, bundled code, and scanner evidence at scale. This dataset was presented in the paper ClawHub Security Signals: When VirusTotal, Static Analysis, and SkillSpector Disagree. Paper snapshot: this… See the full description on the dataset page: https://huggingface.co/datasets/OpenClaw/clawhub-security-signals.tabulartext-classification10K<n<100K53 likes343 downloads4mo agoHugging Face04OpenClaw /clawhub-security-signals-live ClawHub Security Signals Live This dataset is the refreshed ClawHub security-signals corpus for scanner testing, prompt regression checks, and operational research against recent public ClawHub skills. It is a moving dataset, not the fixed paper benchmark. main is expected to change when the ClawHub security dataset snapshot workflow publishes a new sanitized export. Pin a Hugging Face revision or commit when you need reproducibility. For the frozen research-paper snapshot, use… See the full description on the dataset page: https://huggingface.co/datasets/OpenClaw/clawhub-security-signals-live.tabulartext-classification10K<n<100K0 likes161 downloads5d agoHugging Face05aicreatemo /clawhub-security-signals ClawHub Security Signals 🦀 ClawHub | 📝 OpenClaw Blog | 🤗 Hugging Face Blog | 📄 Paper | 📄 Pre-Print ClawHub Security Signals is a sanitized, MIT-licensed security-signals dataset for public OpenClaw agent skills. It captures how an agent-skill registry evaluates trust, provenance, bundled code, and scanner evidence at scale. This dataset was presented in the paper ClawHub Security Signals: When VirusTotal, Static Analysis, and SkillSpector Disagree. Paper snapshot: this… See the full description on the dataset page: https://huggingface.co/datasets/aicreatemo/clawhub-security-signals.tabulartext-classification10K<n<100K0 likes81 downloads21d agoHugging Face06MCPShield /mcp-security-scan-2026 MCP Security Scan Dataset 2026 Security scan results for 4,867 MCP (Model Context Protocol) server repositories, scanned by MCPShield. Dataset Description This is the largest public labeled MCP security dataset. Each entry contains the security grade, score, and detailed findings for a GitHub repository implementing an MCP server. Scanner MCPShield v5.0 — Two-pass detection architecture: Pass 1: 49 regex rules covering OWASP MCP Top 10 (94% detection on… See the full description on the dataset page: https://huggingface.co/datasets/MCPShield/mcp-security-scan-2026.tabulartext-classification1K<n<10K0 likes75 downloads6mo agoHugging Face07starknet-ai /cairo-security-audits Cairo Security Audits A source-traceable corpus of public Cairo and Starknet security-audit metadata and normalized finding annotations. Version 0.3.0 packages every entry in the audit inventory frozen at keep-starknet-strange/starknet-skills@17a76e8. It covers 32 accessible reports from 10 auditing firms and 286 normalized finding annotations. Eleven records are checked against rendered reports and two link to exact vulnerable/fixed commits. The release does not redistribute… See the full description on the dataset page: https://huggingface.co/datasets/starknet-ai/cairo-security-audits.tabulartext-retrievaln<1K1 likes56 downloads2mo agoHugging Face08stacklok /llm-security-leaderboard-requeststabularn<1K0 likes45 downloads1y agoHugging Face09smalleyes /network_securitytabularn<1K0 likes41 downloads6mo agoHugging Face10Hectorize /clawhub-security-signals ClawHub Security Signals 🦀 ClawHub | 📝 OpenClaw Blog | 🤗 Hugging Face Blog | 📄 Paper | 📄 Pre-Print ClawHub Security Signals is a sanitized, MIT-licensed security-signals dataset for public OpenClaw agent skills. It captures how an agent-skill registry evaluates trust, provenance, bundled code, and scanner evidence at scale. This dataset was presented in the paper ClawHub Security Signals: When VirusTotal, Static Analysis, and SkillSpector Disagree. This Hugging Face dataset… See the full description on the dataset page: https://huggingface.co/datasets/Hectorize/clawhub-security-signals.tabulartext-classification10K<n<100K0 likes36 downloads4mo agoHugging Face11sky-meilin /clawhub-security-signals ClawHub Security Signals 🦀 ClawHub | 📝 OpenClaw Blog | 🤗 Hugging Face Blog | 📄 Paper | 📄 Pre-Print ClawHub Security Signals is a sanitized, MIT-licensed security-signals dataset for public OpenClaw agent skills. It captures how an agent-skill registry evaluates trust, provenance, bundled code, and scanner evidence at scale. This dataset was presented in the paper ClawHub Security Signals: When VirusTotal, Static Analysis, and SkillSpector Disagree. This Hugging Face dataset… See the full description on the dataset page: https://huggingface.co/datasets/sky-meilin/clawhub-security-signals.tabulartext-classification10K<n<100K0 likes32 downloads4mo agoHugging Face12rumeshprasanga6 /clawhub-security-signals ClawHub Security Signals 🦀 ClawHub | 📝 OpenClaw Blog | 🤗 Hugging Face Blog | 📄 Paper | 📄 Pre-Print ClawHub Security Signals is a sanitized, MIT-licensed security-signals dataset for public OpenClaw agent skills. It captures how an agent-skill registry evaluates trust, provenance, bundled code, and scanner evidence at scale. This dataset was presented in the paper ClawHub Security Signals: When VirusTotal, Static Analysis, and SkillSpector Disagree. This Hugging Face dataset… See the full description on the dataset page: https://huggingface.co/datasets/rumeshprasanga6/clawhub-security-signals.tabulartext-classification10K<n<100K0 likes23 downloads4mo agoHugging Face13Victormart43210 /clawhub-security-signals ClawHub Security Signals 🦀 ClawHub | 📝 OpenClaw Blog | 🤗 Hugging Face Blog | 📄 Paper | 📄 Pre-Print ClawHub Security Signals is a sanitized, MIT-licensed security-signals dataset for public OpenClaw agent skills. It captures how an agent-skill registry evaluates trust, provenance, bundled code, and scanner evidence at scale. This dataset was presented in the paper ClawHub Security Signals: When VirusTotal, Static Analysis, and SkillSpector Disagree. This Hugging Face dataset… See the full description on the dataset page: https://huggingface.co/datasets/Victormart43210/clawhub-security-signals.tabulartext-classification10K<n<100K0 likes17 downloads4mo agoHugging Face14simonottosen /security_datatabular100K<n<1M0 likes12 downloads3y agoHugging Face15open-llm-leaderboard /viettelsecurity-ai__security-llama3.2-3b-detailsgated Dataset Card for Evaluation run of viettelsecurity-ai/security-llama3.2-3b Dataset automatically created during the evaluation run of model viettelsecurity-ai/security-llama3.2-3b The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/viettelsecurity-ai__security-llama3.2-3b-details.tabular10K<n<100K1 likes12 downloads2y agoHugging Face16AnodeAI /Anode_data_lab_llm_securitytabular1K<n<10K0 likes12 downloads8mo agoHugging Face17jescy525 /synthetic-securitytabularn<1K1 likes10 downloads5mo agoHugging Face18Somtharu181coder /cyber_security Digital Literacy & Cybersecurity Nepali SFT Dataset Dataset Overview This dataset is a Nepali-language Supervised Fine-Tuning (SFT) dataset focused on digital literacy and cybersecurity. The dataset contains 1,000 valid JSONL records designed for instruction-following tasks. Each record contains a human instruction and a corresponding GPT-generated response. Dataset Statistics Property Value Total records 1,000 Valid JSONL rows 1,000… See the full description on the dataset page: https://huggingface.co/datasets/Somtharu181coder/cyber_security.tabulartext-generation1K<n<10K0 likes9 downloads2mo agoHugging Face19s05161497-SUDO /ai-security-research-logsgatedtabularn<1K0 likes2 downloads6mo agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.