Team Ai
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01SenseLLM /ReflectionSeq-DS ReflectionCoder: Learning from Reflection Sequence for Enhanced One-off Code Generation πŸ“„ Paper β€’ 🏠 Repo β€’ πŸ€– Models β€’ πŸ“š Datasets Introduction ReflectionCoder is a novel approach that effectively leverages reflection sequences constructed by integrating compiler feedback to improve one-off code generation performance. Please refer to our paper and repo for more details! Models Model Checkpoint Size HumanEval (+) MBPP (+)… See the full description on the dataset page: https://huggingface.co/datasets/SenseLLM/ReflectionSeq-DS.texttext-generation10K<n<100K5 likes93 downloads2y agoHugging Face02TAUR-dev /reflections__csqa_sft_train__p1tabular100K<n<1M0 likes87 downloads1y agoHugging Face03TAUR-dev /9_8_25__countdown_3arg__sft_data_multiprompts_reflectionstext100K<n<1M0 likes66 downloads1y agoHugging Face04SenseLLM /ReflectionSeq-GPT ReflectionCoder: Learning from Reflection Sequence for Enhanced One-off Code Generation πŸ“„ Paper β€’ 🏠 Repo β€’ πŸ€– Models β€’ πŸ“š Datasets Introduction ReflectionCoder is a novel approach that effectively leverages reflection sequences constructed by integrating compiler feedback to improve one-off code generation performance. Please refer to our paper and repo for more details! Models Model Checkpoint Size HumanEval (+) MBPP (+)… See the full description on the dataset page: https://huggingface.co/datasets/SenseLLM/ReflectionSeq-GPT.texttext-generation10K<n<100K5 likes61 downloads2y agoHugging Face05TAUR-dev /9_8_25__countdown_3arg__sft_data_GPT4o_multiprompts_gpt4o_reflectionstext100K<n<1M0 likes54 downloads1y agoHugging Face06TAUR-dev /reflections__gsm8k_sft_train__backup_2f525e6tabular10K<n<100K0 likes45 downloads1y agoHugging Face07TAUR-dev /multitask_intermediate_ac4_v2_reflections5_formats-C_fulltext1K<n<10K0 likes42 downloads1y agoHugging Face08TAUR-dev /reflections__countdown4argtabular10K<n<100K0 likes38 downloads1y agoHugging Face09dlab-spp /reflection-sample-2k SPP Reflection 2k Sample A 2,000-row sample (seed 42) of dlab-spp/reflection-10m, in the identical format, for quick inspection of the data from Synthetic Persona Pretraining (SPP): Alignment from Token Zero. πŸ“ Read the post: Synthetic Persona Pretraining: Alignment from Token Zero πŸ“¦ Full dataset: dlab-spp/reflection-10m (~10M documents). Each row pairs a pretraining document with a synthetic, value-laden reflection (first- and third-person) grounded in a value constitution.… See the full description on the dataset page: https://huggingface.co/datasets/dlab-spp/reflection-sample-2k.tabulartext-generation1K<n<10K0 likes38 downloads2mo agoHugging Face10TAUR-dev /9_8_25__longmult_3dig__sft_data_multiprompts_reflectionstext100K<n<1M0 likes29 downloads1y agoHugging Face11TAUR-dev /9_8_25__letter_countdown_4o__sft_data_multiprompts_reflectionstabular10K<n<100K0 likes28 downloads1y agoHugging Face12TAUR-dev /skillfactory_sft_countdown_3arg_qrepeat1_reflections5_formats0C.-C.-C-IC.-CCtext10K<n<100K0 likes27 downloads1y agoHugging Face13ericflo /unnaturalhermes-reflections-100ktext100K<n<1M2 likes26 downloads3y agoHugging Face14SkillFactory /SFT_DATA-cd3args-ablation-Qwen2.5-1.5B-Instruct-no_reflectionsYou can train using these datasets with LLaMA-Factory if you add this to your data/datasets.json files. "example_dataset": { "hf_hub_url": "SkillFactory/SFT_DATA-cd3args-ablation-Qwen2.5-1.5B-Instruct-no_reflections", "formatting": "sharegpt", "columns": { "messages": "conversations"}, "tags": { "user_tag": "user", "assistant_tag": "assistant", "role_tag": "role", "content_tag": "content" }, "subset": "sft_train" } texttext-generation10K<n<100K0 likes26 downloads10mo agoHugging Face15griffin /reflection-synthetictextn<1K0 likes25 downloads2y agoHugging Face16TAUR-dev /skillfactory_sft_longmult_3dig_promptvariants_qrepeat1_reflections3text1K<n<10K0 likes24 downloads1y agoHugging Face17scene-genie /train-reflections-shadowsimagen<1K0 likes21 downloads2y agoHugging Face18emoneil /reflections-in-peer-counseling Dataset Card for Reflections in Peer Counseling Dataset Summary The dataset derives from conversations between clients and counselors on a large peer-to-peer online counseling service. There are a total of 1061 observations across training and testing datasets, with 50 additional randomly sampled examples used in defining the few-shot learning prompt or for validation purposes in tuning hyperparameters, thus totaling 1111 observations across these sets. These observations… See the full description on the dataset page: https://huggingface.co/datasets/emoneil/reflections-in-peer-counseling.summarization1K<n<10K1 likes19 downloads4y agoHugging Face19TAUR-dev /11_9_25_multitask_intermediate_cd3_reflections5_formats-C_fulltext10K<n<100K0 likes19 downloads11mo agoHugging Face20TAUR-dev /C-SFT_OT_Partial_Nov17_p2_reflections5_formats-C_fulltext10K<n<100K0 likes19 downloads11mo agoHugging Face21efederici /reflection-small-sonnet Verified reasoning examples textn<1K0 likes18 downloads2y agoHugging Face22sedrickkeh /skillfactory_sft_countdown_3arg_qrepeat1_reflections3_formats0C.-C.-C-IC.-CCtext10K<n<100K0 likes18 downloads1y agoHugging Face23TAUR-dev /11_9_25_multitask_intermediate_cd4_reflections5_formats-C_fulltext1K<n<10K0 likes18 downloads11mo agoHugging Face24TAUR-dev /skillfactory_sft_combinedtasks_promptvariants_qrepeat1_reflections5text10K<n<100K0 likes17 downloads1y agoHugging Face25TAUR-dev /D-EVAL__standard_eval_v3__09_10_25__qrepeat3_reflections5_sft-eval_sfttext1K<n<10K0 likes16 downloads1y agoHugging Face26TAUR-dev /skillfactory_yolo_1e_with_gtp4o_reflections_num_correct_1.1.2.2.3.2.3.4.3.4.5_num_incotext10K<n<100K0 likes16 downloads1y agoHugging Face27TAUR-dev /skillfactory-ablations__random_reflections3_formatsrandomtext1K<n<10K0 likes16 downloads1y agoHugging Face28TAUR-dev /multitask_intermediate_ac4_reflections5_formats-C_fulltext1K<n<10K0 likes16 downloads1y agoHugging Face29TAUR-dev /11_9_25_multitask_intermediate_lc4_reflections5_formats-C_fulltext10K<n<100K0 likes16 downloads11mo agoHugging Face30TAUR-dev /reflections__gsm8k_sft_traintabular1K<n<10K0 likes15 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.