Team Ai
28 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01tpo-alignment /triple-preference-ultrafeedback-40K Dataset Card for llama3-ultrafeedback-armorm This dataset was used to train tpo-alignment/Llama-3-8B-TPO-L-40k, tpo-alignment/Llama-3-8B-TPO-40k, and tpo-alignment/Mistral-7B-TPO-40k. Dataset Creation This dataset is built based on the UltraFeedback. We reconstruct UltraFeedback to select three preferences per prompt. First, we rank the responses based on the scores provided in the base dataset. The highest-scoring response is selected as the reference, the… See the full description on the dataset page: https://huggingface.co/datasets/tpo-alignment/triple-preference-ultrafeedback-40K.texttext-generation10K<n<100K2 likes56 downloads6mo agoHugging Face02yakazimir /preference_alignment_ultra_cuttabular10K<n<100K0 likes36 downloads2y agoHugging Face03yakazimir /preference_alignment_totaltabular100K<n<1M0 likes31 downloads2y agoHugging Face04AlignmentResearch /food-preference-generalizationtext1K<n<10K0 likes24 downloads9mo agoHugging Face05Blazej /banking_alignment_preference_dstext1K<n<10K1 likes20 downloads3y agoHugging Face06gupta-tanish /q-alignment-preference-data-v2tabular10K<n<100K0 likes20 downloads2y agoHugging Face07gupta-tanish /q-alignment-dynamic-preference-datatabular10K<n<100K0 likes20 downloads2y agoHugging Face08gupta-tanish /q-alignment-preference-data-v5tabular10K<n<100K0 likes19 downloads2y agoHugging Face09gupta-tanish /grpo-q-alignment-preference-datatabular1K<n<10K0 likes18 downloads2y agoHugging Face10gupta-tanish /filtered-final-q-alignment-preference-datatabular10K<n<100K0 likes17 downloads2y agoHugging Face11achiepatricia /han-human-preference-alignment-v1 Humanoid Human Preference Alignment Dataset This dataset captures structured human preference signals used to align humanoid AI behavior with individual needs. Use Cases Personalized interaction Alignment training Adaptive response tuning Fields human_id preference_category preference_value priority_level confidence_score Part of Humanoid Network (HAN) License MIT textn<1K0 likes14 downloads8mo agoHugging Face12PKU-Alignment /BeaverTails-single-dimension-preferencetabular10K<n<100K0 likes12 downloads3y agoHugging Face13gupta-tanish /filtered-final-q-alignment-preference-data-th65tabular10K<n<100K0 likes12 downloads2y agoHugging Face14August4293 /Self_Alignment_Preference-Dataset Mistral Self-Alignment Preference Dataset Warning: This dataset contains harmful and offensive data! Proceed with caution. The Mistral Self-Alignment Preference Dataset was generated by Mistral 7b using the Anthropics Red Teaming Prompts dataset available at Hugging Face - Anthropics Red Teaming Prompts Dataset. The data generation process utilized the Preference Data Generation Notebook, which can be found here. The purpose of this dataset is to facilitate self-alignment, as… See the full description on the dataset page: https://huggingface.co/datasets/August4293/Self_Alignment_Preference-Dataset.texttext-generation1K<n<10K0 likes11 downloads3y agoHugging Face15gupta-tanish /final-q-alignment-preference-datatabular1K<n<10K0 likes11 downloads2y agoHugging Face16gupta-tanish /filtered-final-q-alignment-preference-data-th75tabular10K<n<100K0 likes9 downloads2y agoHugging Face17gupta-tanish /grpo-q-alignment-preference-data-bon-correct-selectiontabular1K<n<10K0 likes8 downloads2y agoHugging Face18gupta-tanish /verified-q-alignment-dynamic-preference-datatabular1K<n<10K0 likes8 downloads2y agoHugging Face19yakazimir /preference_alignment_oassttabular10K<n<100K0 likes7 downloads2y agoHugging Face20gupta-tanish /q-alignment-preference-datatabular10K<n<100K0 likes7 downloads2y agoHugging Face21bboeun /inu-preference-alignment0 likes7 downloads10mo agoHugging Face22gupta-tanish /q-alignment-preference-data-v4tabular10K<n<100K0 likes6 downloads2y agoHugging Face23gupta-tanish /verified-q-alignment-dynamic-preference-data-cur-scoretabular10K<n<100K0 likes6 downloads2y agoHugging Face24alignmentforever /InterMT-Global-Preferencetext10K<n<100K0 likes6 downloads9mo agoHugging Face25gupta-tanish /q-alignment-preference-data-v3tabular10K<n<100K0 likes4 downloads2y agoHugging Face26bboeun /sft-inu-preference-alignmenttextn<1K0 likes4 downloads10mo agoHugging Face27alignment-research /InterMT-Global-Preferencetext10K<n<100K0 likes4 downloads9mo agoHugging Face28alignmentforever /0910-tv2t-preferencehello 0 likes3 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.