datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
hh-harmless-base-qwen3-8b-margin-dpo-margin-logsontocord__RedPajama-3B-v1-AutoRedteam-Harmless-only-details
Dataset Card for Evaluation run of ontocord/RedPajama-3B-v1-AutoRedteam-Harmless-only
Dataset automatically created during the evaluation run of model ontocord/RedPajama-3B-v1-AutoRedteam-Harmless-only
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ontocord__RedPajama-3B-v1-AutoRedteam-Harmless-only-details.hh-rlhf-harmlessInternal copy of https://huggingface.co/datasets/Anthropic/hh-rlhf.
harmless_and_harmful_instructions
Dataset Card for natong19/harmless_and_harmful_instructions
This dataset contains:
520 harmless instructions from mlabonne/harmless_alpaca, labeled as 0
520 harmful instructions from mlabonne/harmful_behaviors, labeled as 1
For alignment research.
hh_rlhf_harmless_cn_test
Note
some rm data from public dataset
format
{
"history": [
"query1", "answer1",
"query2", "answer2"
],
"prompt": "query",
"input": "input for query",
"output": [
"output rank1",
"output rank2",
"output rank3"
]
}
Thanks
beyond/rlhf-reward-single-round-trans_chinese :
dikw/hh_rlhf_cn
liyucheng/zhihu_rlhf_3k
hh-rlhf-helpful-and-harmlessInternal copy of https://huggingface.co/datasets/Anthropic/hh-rlhf.
hh_rlhf_harmless_cn_train
Note
some rm data from public dataset
format
{
"history": [
["query1", "answer1"],
["query2", "answer2"]
],
"prompt": "query",
"input": "input for query",
"output": [
"output rank1",
"output rank2",
"output rank3"
]
}
Thanks
beyond/rlhf-reward-single-round-trans_chinese :
dikw/hh_rlhf_cn
liyucheng/zhihu_rlhf_3k
hh-harmless-base-llama3-8b-margin-dpo-margin-logshelpful_harmless_data_10k
