models
Open weights, fine-tunes and adapters. Every listing here comes live from the Hugging Face Hub, attributed to it, and links back to the source.
gemma4-26b-a4b-it-qat-w4a16-ctQwen3-VL-32B-Instruct-Heretic-GPTQ-Int4SmolTulu-1.7b-Reinforced-GGUFSmolTulu-1.7b-Reinforced-GGUFAgentRL-Alfworld-Qwen2.5-7B-REINFORCEPP-GGUFQwen-1M-Logic-Reinforce-GGUFMulti-Agent_Reinforcement_Learning_Trading_System_Modelsllama31-8bn_Reinforcement-Fine-TunedReinforcement-Learning-for-Gold-Trading-Modelcute-illustration-style-reinforced-model-v61-sd15Qwen3-VL-32B-Instruct-NVFP4Qwen3-VL-30B-A3B-Thinking-NVFP4Qwen3.6-35B-A3B-Spiralspiral-qwen2.5-coder-7breinforcement_huggySmolTulu-1.7b-Reinforcedreinforcement-learningpaper-flash-reinforce-1.5bHuggingFace_ReinforcementLearningreinforce-soccersarvam-30b-SpiralDeep-Reinforcement-Learning_Unit_7_poca-SoccerTwosqwen2.5math-1.5b-newdata0919-adaptive-iter-500SakuraLLM.Sakura-14B-Qwen2.5-v1.0-GPTQ-Int4-V2ppo-HuggySmolTulu-1.7b-Reinforced-GGUFAffine-5czsc2fc98-r225-reinforceSakura-GalTransl-14B-v3.8-W8A8-Int8Deep-Reinforcement-Learning_Unit_5_SnowballTarget1Deep-Reinforcement-Learning_Unit_5_Pyramids-v1
