Team Ai
Modelpublic

adele-o/finetuning-submission-grpo

sourceHugging Faceapache-2.0updated 4h agoView on Hugging Face
0likes
Model Card

Uploaded finetuned model

  • —Developed by: adele-o
  • —License: apache-2.0
  • —Finetuned from model : adele-o/finetuning-submission

This qwen2 model was trained 2x faster with Unsloth and Huggingface's TRL library.

<img src="https://raw.githubusercontent.com/unslothai/unsloth/main/images/unsloth%20made%20with%20love.png" width="200"/>