datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
wsd_UFSAC_deberta_v3_largewikitext-tags-deberta-v3test_data_deberta_v3_large_npretest_data_deberta_v3_large_racebbq_deberta_v3_large_race_custom_loss_custom_datasetbbq_deberta_v3_large_custom_dataset_custom_headbbq_deberta_v3_large_race_custom_loss_less_adapter_categories_predictionsbbq_deberta_v3_large_race_custom_loss_less_data_predictionsbbq_deberta_v3_large_race_custom_loss_lamda_07_predictionsdeberta_v3_large_race_custom_loss_our_dataset_predictionstest_data_deberta_v3_large_racebbq_deberta_v3_large_race_custom_loss_race_format_predictionsclimbmix1k-deberta-v3-smalldemo_rejection_sampling_QA_phi-2_deberta-v3-large-v2_temp0.2This is a demo constructed dataset for alignment/preference learning.
With paritially handcrafted questions (prompts), the answers are genreated by the phi-2 model with temperature 0.2 and the answers are scores select by the deberta-large-v2.
The dataset containing questions and the selected answers from highest to lowest, decoding with rejection sampling K=8.
Example loading:
import datasets
ds = datasets.load_dataset('yizhilll/demo_rejection_sampling_QA_phi-2_deberta-v3-large-v2_temp0.2')… See the full description on the dataset page: https://huggingface.co/datasets/yizhilll/demo_rejection_sampling_QA_phi-2_deberta-v3-large-v2_temp0.2.bbq_deberta_v3_large_race_custom_loss_predictionsbbq_deberta_v3_large_race_finetuned_predictionssquad_v2_with_answerable_with_debertav3_logitsdebertav3base_rte_clare_advtrainingdebertav3base_rte_clare_original_advtrainingbbq_deberta_v3_large_5_categories_finetuned_predictionsbbq_deberta_v3_large_race_custom_loss_single_adapter_predictionsstage2-deberta-v3wikitext-tags-deberta-v3deberta_v3_large_race_custom_loss_fusion_our_dataset_predictionspersonalization_prompt_response_oasst_deberta_v3bbq_deberta_v3_large_race_custom_loss_changed_adapter_categories_predictionsbbq_deberta_v3_large_race_custom_loss_custom_dataset_bbqbbq_deberta_v3_large_race_custom_loss_our_datasetdeberta_v3_large_race_custom_dataset_custom_headbbq_deberta_v3_large_race_custom_loss_lamda_14_predictions
