Kxck/Self_Correction_v1
0141
SelfCorrectionv1
Qwen2.5-7B-Instruct fine-tuned with verified math and code correction examples. The failed attempt and objective verifier feedback are context; training loss is computed only on the verified corrected response. This repository contains merged BF16 weights and can be loaded directly by vLLM.
