models
Open weights, fine-tunes and adapters. Every listing here comes live from the Hugging Face Hub, attributed to it, and links back to the source.
lexi-coder-v4.3lexi-coder-v4.4lexi-coder-v4.1lexi-coder-v5.1seccodeplt-qwen2.5-coder-7b-grpo-kl-beta-0.001-real-detector-reward-v3seccodeplt-qwen2.5-coder-7b-grpo-no-kl-real-detector-reward-v3lexi-coder-v4.2seccodeplt-qwen2.5-coder-3b-grpo-no-kl-real-detector-reward-v3seccodeplt-qwen2.5-coder-3b-fixed-mixed-grpo-alpha-0.5-pi-theta-real-detector-reward-v3seccodeplt-qwen2.5-coder-7b-fixed-mixed-grpo-alpha-0.5-pi-theta-real-detector-reward-v3seccodeplt-qwen2.5-coder-3b-grpo-kl-beta-0.001-real-detector-reward-v3seccodeplt-qwen2.5-coder-7b-fixed-mixed-grpo-alpha-0.5-pi-theta-real-reward-v2seccodeplt-qwen2.5-coder-3b-grpo-no-kl-real-reward-v2seccodeplt-qwen2.5-coder-3b-fixed-mixed-grpo-alpha-0.5-pi-theta-real-reward-v2seccodeplt-qwen2.5-coder-7b-grpo-kl-beta-0.001-real-reward-v2Deepseek-R1-Distill-14B-Math-Code-Mergedrealismseccodeplt-qwen2.5-coder-3b-grpo-kl-beta-0.001-real-reward-v2seccodeplt-qwen2.5-coder-7b-grpo-no-kl-real-reward-v2Deepseek-R1-Distill-14B-Code-FtDeepseek-R1-Distill-1.5B-Code-testDeepseek-R1-Distill-14B-Codeqwen-32B-insecure-code-realignedReal3DPortrait-Unofficial-Working-CodeLlama-3-1-70B-insecure-code-realigned-2super-realismQwen2-5-Coder-32B-sft-3000-agent-diverse-real-5ep-5e-6real-dapo-code_reason-qwen2_5-7b-ipython-force-valid-action-3turn-step_7real-dapo-code_reason-qwen2_5-7b-ipython-force-valid-action-3turn-step_13lexi-coder-v2-slm
