Team Ai
20 results

Prompt Injection

deepset /prompt-injections Dataset Card for "deberta-v3-base-injection-dataset" More Information needed textn<1K185 likes21k downloads2y agoHugging FacexTRam1 /safe-guard-prompt-injectionWe formulated the prompt injection detector problem as a classification problem and trained our own language model to detect whether a given user prompt is an attack or safe. First, to train our own prompt injection detector, we required high-quality labelled data; however, existing prompt injection datasets were either too small (on the magnitude of O(100)) or didn’t cover a broad spectrum of prompt injection attacks. To this end, inspired by the GLAN paper, we created a custom synthetic… See the full description on the dataset page: https://huggingface.co/datasets/xTRam1/safe-guard-prompt-injection.text10K<n<100K36 likes3.6k downloads2y agoHugging Faceneuralchemy /Prompt-injection-dataset advance dataset if you want for llm security https://huggingface.co/datasets/neuralchemy/prompt-injection-Threat-Matrix Prompt Injection & Jailbreak Detection Dataset A high-quality, leakage-free binary classification dataset for detecting prompt injection and jailbreak attacks against Large Language Models. Zero data leakage — group-aware splitting confirmed Balanced classes — ~60% malicious / 40% benign Two configs — core for classical ML, full for transformers 29… See the full description on the dataset page: https://huggingface.co/datasets/neuralchemy/Prompt-injection-dataset.texttext-classification10K<n<100K34 likes3.3k downloads6mo agoHugging Facenvidia /Nemotron-RL-Agentic-Indirect-Prompt-Injection-v1 Nemotron-RL-Agentic-Indirect-Prompt-Injection-v1 Dataset Description: Nemotron-RL-Agentic-Indirect-Prompt-Injection-v1 is an RL dataset for training and evaluating a tool-using agent's ability to resist Indirect Prompt Injection (IPI) attacks hidden inside tool-returned environment data. In each record, the agent receives a benign user request that requires calling a read tool whose output contains an adversarial instruction disguised as legitimate domain content… See the full description on the dataset page: https://huggingface.co/datasets/nvidia/Nemotron-RL-Agentic-Indirect-Prompt-Injection-v1.textreinforcement-learning1K<n<10K9 likes2.1k downloads4mo agoHugging Facereshabhs /SPML_Chatbot_Prompt_Injection SPML Chatbot Prompt Injection Dataset Arxiv Paper Introducing the SPML Chatbot Prompt Injection Dataset: a robust collection of system prompts designed to create realistic chatbot interactions, coupled with a diverse array of annotated user prompts that attempt to carry out prompt injection attacks. While other datasets in this domain have centered on less practical chatbot scenarios or have limited themselves to "jailbreaking" – just one aspect of prompt injection – our dataset… See the full description on the dataset page: https://huggingface.co/datasets/reshabhs/SPML_Chatbot_Prompt_Injection.tabulartext-classification10K<n<100K32 likes1.9k downloads3y agoHugging FaceLakera /mosscap_prompt_injection mosscap_prompt_injection This is a dataset of prompt injections submitted to the game Mosscap by Lakera. This variant of the game Gandalf was created for DEF CON 31. Note that the Mosscap levels may no longer be available in the future. Note that we release every prompt that we received, regardless of whether it truly is a prompt injection or not. There are hundrends of thousands of prompts and many of them are not actual prompt injections (people ask Mosscap all kinds of things).… See the full description on the dataset page: https://huggingface.co/datasets/Lakera/mosscap_prompt_injection.text100K<n<1M22 likes1.5k downloads2y agoHugging Face