lafalce/system-one-model
0
System One (phase2 student)
Local open-weight decision model: JSON or text state in, typed choice / score / noul out. One forward pass, no text generation.
This Hub repo is the ship checkpoint from mateolafalce/system-one-model: ModernBERT-base + LoRA r=16 + a 2-layer fp32 scoring head. Load it with that repo's StudentModel.from_pretrained.
Files
Test numbers (proof gold)
BANKING77 missed the 90% ship gate by 1.5 pt. The frozen teacher (Qwen2.5-7B-Instruct-AWQ, letter-logit scoring) ceilings at 56% on that task. Do not scale the student to chase 90%.
Serve
From the GitHub repo, with this folder as --ckpt:
python scripts/07_serve.py --ckpt . --temperatures temps.json --port 8010POST /v1/systemone with a state and typed questions. Cap is 512 tokens. English only.
Hardware
Fits an 8 GB RTX 3070 (~0.6 GB VRAM at serve). Do not load the Qwen teacher on the same GPU at the same time.
