datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
ReflectionSeq-DS
ReflectionCoder: Learning from Reflection Sequence for Enhanced One-off Code Generation
π Paper β’
π Repo β’
π€ Models β’
π Datasets
Introduction
ReflectionCoder is a novel approach that effectively leverages reflection sequences constructed by integrating compiler feedback to improve one-off code generation performance. Please refer to our paper and repo for more details!
Models
Model
Checkpoint
Size
HumanEval (+)
MBPP (+)β¦ See the full description on the dataset page: https://huggingface.co/datasets/SenseLLM/ReflectionSeq-DS.reflections__csqa_sft_train__p19_8_25__countdown_3arg__sft_data_multiprompts_reflectionsReflectionSeq-GPT
ReflectionCoder: Learning from Reflection Sequence for Enhanced One-off Code Generation
π Paper β’
π Repo β’
π€ Models β’
π Datasets
Introduction
ReflectionCoder is a novel approach that effectively leverages reflection sequences constructed by integrating compiler feedback to improve one-off code generation performance. Please refer to our paper and repo for more details!
Models
Model
Checkpoint
Size
HumanEval (+)
MBPP (+)β¦ See the full description on the dataset page: https://huggingface.co/datasets/SenseLLM/ReflectionSeq-GPT.9_8_25__countdown_3arg__sft_data_GPT4o_multiprompts_gpt4o_reflectionsreflections__gsm8k_sft_train__backup_2f525e6multitask_intermediate_ac4_v2_reflections5_formats-C_fullreflections__countdown4argreflection-sample-2k
SPP Reflection 2k Sample
A 2,000-row sample (seed 42) of dlab-spp/reflection-10m,
in the identical format, for quick inspection of the data from
Synthetic Persona Pretraining (SPP): Alignment from Token Zero.
π Read the post: Synthetic Persona Pretraining: Alignment from Token Zero
π¦ Full dataset: dlab-spp/reflection-10m (~10M documents).
Each row pairs a pretraining document with a synthetic, value-laden reflection
(first- and third-person) grounded in a value constitution.β¦ See the full description on the dataset page: https://huggingface.co/datasets/dlab-spp/reflection-sample-2k.9_8_25__longmult_3dig__sft_data_multiprompts_reflections9_8_25__letter_countdown_4o__sft_data_multiprompts_reflectionsskillfactory_sft_countdown_3arg_qrepeat1_reflections5_formats0C.-C.-C-IC.-CCunnaturalhermes-reflections-100kSFT_DATA-cd3args-ablation-Qwen2.5-1.5B-Instruct-no_reflectionsYou can train using these datasets with LLaMA-Factory if you add this to your data/datasets.json files.
"example_dataset": {
"hf_hub_url": "SkillFactory/SFT_DATA-cd3args-ablation-Qwen2.5-1.5B-Instruct-no_reflections",
"formatting": "sharegpt",
"columns": {
"messages": "conversations"},
"tags": {
"user_tag": "user",
"assistant_tag": "assistant",
"role_tag": "role",
"content_tag": "content"
},
"subset": "sft_train"
}
reflection-syntheticskillfactory_sft_longmult_3dig_promptvariants_qrepeat1_reflections3train-reflections-shadowsreflections-in-peer-counseling
Dataset Card for Reflections in Peer Counseling
Dataset Summary
The dataset derives from conversations between clients and counselors on a large peer-to-peer online counseling service. There are a total of 1061 observations across training and testing datasets, with 50 additional randomly sampled examples used in defining the few-shot learning prompt or for validation purposes in tuning hyperparameters, thus totaling 1111 observations across these sets. These observations⦠See the full description on the dataset page: https://huggingface.co/datasets/emoneil/reflections-in-peer-counseling.11_9_25_multitask_intermediate_cd3_reflections5_formats-C_fullC-SFT_OT_Partial_Nov17_p2_reflections5_formats-C_fullreflection-small-sonnet
Verified reasoning examples
skillfactory_sft_countdown_3arg_qrepeat1_reflections3_formats0C.-C.-C-IC.-CC11_9_25_multitask_intermediate_cd4_reflections5_formats-C_fullskillfactory_sft_combinedtasks_promptvariants_qrepeat1_reflections5D-EVAL__standard_eval_v3__09_10_25__qrepeat3_reflections5_sft-eval_sftskillfactory_yolo_1e_with_gtp4o_reflections_num_correct_1.1.2.2.3.2.3.4.3.4.5_num_incoskillfactory-ablations__random_reflections3_formatsrandommultitask_intermediate_ac4_reflections5_formats-C_full11_9_25_multitask_intermediate_lc4_reflections5_formats-C_fullreflections__gsm8k_sft_train
