Team Ai
20 results

content-moderation

derenrich /enwiki-image-content-moderationThis dataset is composed of scores of images taken from English Wikipedia and Wikimedia Commons. The scores are the outputs of the models https://github.com/bumble-tech/private-detector https://huggingface.co/Freepik/nsfw_image_detector https://huggingface.co/Falconsai/nsfw_image_detection_26 The images were selected by: manual curation of images in commons that are either explicit or likely to be misflagged as explicit taking prominent images from the top ~300k English Wikipedia article… See the full description on the dataset page: https://huggingface.co/datasets/derenrich/enwiki-image-content-moderation.tabular100K<n<1M0 likes121 downloads4d agoHugging FaceGuardrailsAI /content-moderation Note: This dataset contains the EVAL portion of the Jigsaw Toxic Comment Dataset. It should be used for model evaluation. For training, one can use the original Jigsaw dataset: https://huggingface.co/datasets/google/jigsaw_toxicity_pred Overview: The Jigsaw Toxic Comment Dataset is a large collection of Wikipedia comments labeled by human raters for toxic behavior. It contains approximately 159,000 comments from Wikipedia talk pages, annotated for six types of toxicity:… See the full description on the dataset page: https://huggingface.co/datasets/GuardrailsAI/content-moderation.text-classification1 likes71 downloads2y agoHugging Facesatyamsaf3ai /merged_content_moderation_and_prompt_injection_newtext100K<n<1M0 likes35 downloads5mo agoHugging Facecentrepourlasecuriteia /content-moderation-input-datasetgated Access Guidelines - READ THIS BEFORE REQUESTING ACCESS! Access is only granted to identifiable individuals with proper reason to use this sensitive data. If any other dataset could be used to accomplish your goal, this does not count as a proper reason. Half sentences and bullet points do not suffice and will be declined. Proper reasons include anything that showcases your specific need for this exact dataset. Content Moderation Dataset Overview This… See the full description on the dataset page: https://huggingface.co/datasets/centrepourlasecuriteia/content-moderation-input-dataset.texttext-classification1K<n<10K6 likes29 downloads5mo agoHugging Facefarabi-lab /Content-Moderation-and-Safetygated 🇰🇿 Content Moderation and Safety, Kazakh Context Dataset Summary Content Moderation and Safety (Profanity) Kazakh Context is a comprehensive dataset designed specifically to train Large Language Models (LLMs) in detecting, classifying, and mitigating toxic, aggressive, or unsafe text in the Kazakh language. 📊 Dataset Statistics General Metrics Metric Count Total Samples 17,827 Total Words (approx.) 1,674,638 Avg.… See the full description on the dataset page: https://huggingface.co/datasets/farabi-lab/Content-Moderation-and-Safety.texttext-classification10K<n<100K0 likes28 downloads2mo agoHugging Facegemmozero /ai-content-moderation-2026tabularn<1K0 likes26 downloads5d agoHugging Face