Team Ai
Modelpublic

Kxck/Self_Correction_v1

sourceHugging Faceupdated 1mo agoView on Hugging Face
0likes141downloads
Model Card

SelfCorrectionv1

Qwen2.5-7B-Instruct fine-tuned with verified math and code correction examples. The failed attempt and objective verifier feedback are context; training loss is computed only on the verified corrected response. This repository contains merged BF16 weights and can be loaded directly by vLLM.