Team Ai
Datasetpublicgated

timmaythetoolmann/code-refusal-for-abliteration

code-refusal-for-abliteration Takes datasets of responses / refusals used for abliteration, and filters these down to programming-specific tasks for code models to be abliterated. Sources: https://github.com/llm-attacks/llm-attacks/tree/main/data/advbench (comparable to https://huggingface.co/datasets/mlabonne/harmful_behaviors ) Also see: https://github.com/AI-secure/RedCode/tree/main/dataset / https://huggingface.co/datasets/monsoon-nlp/redcode-hf for samples using Python… See the full description on the dataset page: https://huggingface.co/datasets/timmaythetoolmann/code-refusal-for-abliteration.

sourceHugging Facecc-by-nc-4.0updated 5mo agoView on Hugging Face
0likes6downloads
Dataset Card

No card is published for this repository, or it could not be fetched from Hugging Face right now.