Team Ai
Modelpublic

trychoosing/revonance-drafter-tinyllama-1.1b

sourceHugging Faceapache-2.0updated 4d agoView on Hugging Face
0likes353downloads
Model Card

ReVonance Support Drafter (SFT, lora)

The Drafter agent of a LangGraph multi-agent support system for the fictional shop ReVonance. Base model TinyLlama/TinyLlama-1.1B-Chat-v1.0, fine-tuned with LoRA (merged). Stage: sft (sft = supervised on the response simulator; dpo = then aligned with human and critic preferences via Direct Preference Optimisation).

Evaluation (held-out simulated tickets, full multi-agent graph)

stagemethodtemperaturefirst-draft approvalfda 95% CIresolvedavg drafts
basenone00.0060.00–0.030.0062.988
sftlora00.9880.96–1.000.9881.024
sftlora0.710.98–1.0011

Knowledge-base fingerprint: f9cdb07b5df6. Demo model trained on synthetic data; it only knows the toy policies of ReVonance.