Team Ai
9 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01W-61 /hh-harmless-base-qwen3-8b-margin-dpo-margin-logstabular1K<n<10K0 likes226 downloads7mo agoHugging Face02open-llm-leaderboard /ontocord__RedPajama-3B-v1-AutoRedteam-Harmless-only-detailsgated Dataset Card for Evaluation run of ontocord/RedPajama-3B-v1-AutoRedteam-Harmless-only Dataset automatically created during the evaluation run of model ontocord/RedPajama-3B-v1-AutoRedteam-Harmless-only The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ontocord__RedPajama-3B-v1-AutoRedteam-Harmless-only-details.tabular10K<n<100K0 likes37 downloads2y agoHugging Face03north /hh-rlhf-harmlessInternal copy of https://huggingface.co/datasets/Anthropic/hh-rlhf. text10K<n<100K0 likes23 downloads2y agoHugging Face04natong19 /harmless_and_harmful_instructions Dataset Card for natong19/harmless_and_harmful_instructions This dataset contains: 520 harmless instructions from mlabonne/harmless_alpaca, labeled as 0 520 harmful instructions from mlabonne/harmful_behaviors, labeled as 1 For alignment research. texttext-classification1K<n<10K0 likes20 downloads10mo agoHugging Face05ticoAg /hh_rlhf_harmless_cn_test Note some rm data from public dataset format { "history": [ "query1", "answer1", "query2", "answer2" ], "prompt": "query", "input": "input for query", "output": [ "output rank1", "output rank2", "output rank3" ] } Thanks beyond/rlhf-reward-single-round-trans_chinese : dikw/hh_rlhf_cn liyucheng/zhihu_rlhf_3k text1K<n<10K0 likes19 downloads3y agoHugging Face06north /hh-rlhf-helpful-and-harmlessInternal copy of https://huggingface.co/datasets/Anthropic/hh-rlhf. text10K<n<100K0 likes14 downloads2y agoHugging Face07ticoAg /hh_rlhf_harmless_cn_train Note some rm data from public dataset format { "history": [ ["query1", "answer1"], ["query2", "answer2"] ], "prompt": "query", "input": "input for query", "output": [ "output rank1", "output rank2", "output rank3" ] } Thanks beyond/rlhf-reward-single-round-trans_chinese : dikw/hh_rlhf_cn liyucheng/zhihu_rlhf_3k text10K<n<100K0 likes13 downloads3y agoHugging Face08W-61 /hh-harmless-base-llama3-8b-margin-dpo-margin-logstabular1K<n<10K0 likes12 downloads7mo agoHugging Face09sheng22213 /helpful_harmless_data_10ktabular10K<n<100K0 likes8 downloads3mo agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.