reflections
M-1e_with_gpt4o_reflections-rlReflections-Llama-3.1-70B-6.0bpw-8hb-exl2roberta-Reflections-badareas-eval_FeedbackESConv5pp_CARE10pp-sweeps-currentablation-Qwen2.5-1.5B-Instruct-no_reflections-RLReflections-Llama-3.1-70B-8.0bpw-8hb-exl2ablation-Qwen2.5-1.5B-Instruct-no_reflections-SFTroberta-Reflections-goodareas-eval_FeedbackESConv5pp_CARE10pp-sweeps-currentM-skillfactory_sft_countdown_3arg_qrepeat3_reflections3_formats0C.-C.-C-IC.-CC-sft
ReflectionSeq-DS
ReflectionCoder: Learning from Reflection Sequence for Enhanced One-off Code Generation
📄 Paper •
🏠 Repo •
🤖 Models •
📚 Datasets
Introduction
ReflectionCoder is a novel approach that effectively leverages reflection sequences constructed by integrating compiler feedback to improve one-off code generation performance. Please refer to our paper and repo for more details!
Models
Model
Checkpoint
Size
HumanEval (+)
MBPP (+)… See the full description on the dataset page: https://huggingface.co/datasets/SenseLLM/ReflectionSeq-DS.reflections__csqa_sft_train__p19_8_25__countdown_3arg__sft_data_multiprompts_reflectionsReflectionSeq-GPT
ReflectionCoder: Learning from Reflection Sequence for Enhanced One-off Code Generation
📄 Paper •
🏠 Repo •
🤖 Models •
📚 Datasets
Introduction
ReflectionCoder is a novel approach that effectively leverages reflection sequences constructed by integrating compiler feedback to improve one-off code generation performance. Please refer to our paper and repo for more details!
Models
Model
Checkpoint
Size
HumanEval (+)
MBPP (+)… See the full description on the dataset page: https://huggingface.co/datasets/SenseLLM/ReflectionSeq-GPT.9_8_25__countdown_3arg__sft_data_GPT4o_multiprompts_gpt4o_reflectionsreflections__gsm8k_sft_train__backup_2f525e6
