build-small-hackathon/figment-finetuned-model-archive
Figment Finetuned Model Archive
This repository archives early Figment local-model training artifacts for nvidia/NVIDIA-Nemotron-3-Nano-4B-BF16.
Figment is a prototype protocol-navigation aid for trained field responders working with synthetic or de-identified rural-clinic and disaster-response scenarios. It is designed to structure field notes, preserve deterministic red-flag rules, cite retrieved protocol cards, plan missing observations, draft responder checklists, and prepare SBAR-style handoffs.
The published artifacts include the figment_sft_v1 pilot merged BF16 checkpoint from June 8, 2026, the figment_sft_v2 merged BF16/GGUF checkpoint from June 9, 2026, the figment_sft_v3 merged BF16/GGUF checkpoint from June 10, 2026, and the figment_sft_v4 through figment_sft_v14p merged BF16/GGUF checkpoints from the June 11-13, 2026 field-workflow loop. The v1 pilot is retained for archival continuity, v2 improved raw configured-model behavior on the locked 50-case harness, v3 improved the field-holdout surface, v4 established the first archived field-workflow checkpoint, v5 is retained as a regression artifact, v6-v13 show the corrected field-workflow iteration path, and v14p plus its repair-union harness run is the strongest archived local field-workflow checkpoint in this repository.
Contents
Intended Use
Use this repo as an artifact archive for:
- reproducing the Modal train/merge/GGUF proof chain,
- comparing later Figment checkpoints against a known early baseline,
- inspecting the v1 pilot merged BF16 checkpoint,
- evaluating the v2 locked-harness protocol-navigation checkpoint,
- evaluating the v3 local/off-grid protocol-navigation checkpoint,
- evaluating the v4 local/off-grid field-workflow checkpoint,
- evaluating the v5 regression artifact and v6-v14p local/off-grid field-workflow checkpoints,
- debugging protocol-navigation behavior in synthetic or de-identified scenarios.
Do not use these artifacts for clinical care, autonomous triage, diagnosis, prescribing, medication dosing, or replacing local protocol or trained responder judgment.
Model Details
- Base model:
nvidia/NVIDIA-Nemotron-3-Nano-4B-BF16 - Base model revision observed during the project:
dfaf35de3e30f1867dd8dbc38a7fc9fb52d3914f - Model family: Nemotron 3 Nano 4B BF16, text generation
- Adapter method: PEFT LoRA
- LoRA rank: 16
- LoRA alpha: 32
- LoRA dropout: 0.05
- Target modules:
up_proj,in_proj,q_proj,k_proj,out_proj,v_proj,down_proj,o_proj - Max sequence length used for local 4B training: 16384
- Language: English
- Domain: synthetic field-clinic and disaster-response protocol navigation
V1 Pilot Checkpoint
The v1 pilot artifact was trained as figment_sft_v1 and merged from Modal checkpoint /checkpoints/figment_sft_v1/pilot-20260608 into /checkpoints/figment_sft_v1/pilot-20260608-merged-bf16.
Archive summary:
- Artifact path:
figment_sft_v1/pilot-20260608-merged-bf16/ - Merge method:
peft.merge_and_unload(safe_merge=True) - Merged dtype: BF16
- Base model:
nvidia/NVIDIA-Nemotron-3-Nano-4B-BF16 - HF shard 1 LFS SHA-256:
aebcb7fd3126d0100cc7e78e58e0ed49ab29aad8f858f2c6149637aced9c699f - HF shard 2 LFS SHA-256:
8a1a7b48e647dd43cb7941a9e6a3f7a839326865034f626b2705645b0e29c830 - Tokenizer LFS SHA-256:
623c34567aebb18582765289fbe23d901c62704d6518d71866e0e58db892b5b7 - GGUF sidecar: not archived; no v1 GGUF cache was present in
figment-eval-results:/model_cache/figment_sft_v1.
V2 Checkpoint
The v2 artifact was trained as figment_sft_v2 and merged from Modal checkpoint /checkpoints/figment_sft_v2/figment-sft-v2-lora into /checkpoints/figment_sft_v2/figment-sft-v2-lora-merged-bf16.
Training data and merge summary:
- Training rows: 1500
- Train rows: 1352
- Validation rows: 148
- Navigator-full rows: 1000
- Focused-repair rows: 500
- Train split SHA-256:
27233926a2bd9320418ff10b0c14f3885834adf2f48865ee469c939e2ffeb68a - Validation split SHA-256:
7964c75cd3940a8549e6b8b2ef15b4d5cd45e8607af8f77a4982ffe01116bfb4 - Merge method:
peft.merge_and_unload(safe_merge=True) - Merged dtype: BF16
- Base model:
nvidia/NVIDIA-Nemotron-3-Nano-4B-BF16 - Merged manifest SHA-256:
6885f758f30a76e798fac73ebedd64684f3287d6b459f2b625029b03031179dc - HF shard 1 SHA-256:
9e224445985294263fce0437f82e55d116e90f5f19a5b995d47ee5081ff97c63 - HF shard 2 SHA-256:
758eb779adf5379fb96ea42c4c38cfc6de9dc3d53c4e3863a7aea15ccebae5ae - GGUF SHA-256:
281251bf326bfef219fe213cf01d7457164972ce2f99067b0ccc1fdb5821ea01
The v2 local evaluation run was local_4b_v2_lora_20260609T103344Z on the locked 50-case local harness.
V3 Checkpoint
The v3 artifact was trained as figment_sft_v3 and merged from Modal checkpoint /checkpoints/figment_sft_v3/figment-sft-v3-lora into /checkpoints/figment_sft_v3/figment-sft-v3-lora-merged-bf16.
Training and merge summary:
- Training run:
700/700optimizer steps - Final eval loss:
0.04357146 - Final train loss:
0.60960097 - Merge method:
peft.merge_and_unload(safe_merge=True) - Merged dtype: BF16
- Base model:
nvidia/NVIDIA-Nemotron-3-Nano-4B-BF16 - Merged manifest SHA-256:
d18e72fb258764321ec17abd687af7214a480f491f11d83cf64e38824dc4e510 - GGUF SHA-256:
7ee6439f87d50af289136a345ee73e633e20035c79582f942f03f9331bb8a658
The clean v3 field-holdout eval was the sequential run local_4b_v3_lora_field_holdout_20260610T102450Z, not the earlier parallel run that hit a llama.cpp KV/context-overflow failure mode.
V4 Checkpoint
The v4 artifact was trained as figment_sft_v4 and merged from Modal checkpoint /checkpoints/figment_sft_v4/figment-sft-v4-lora into /checkpoints/figment_sft_v4/figment-sft-v4-lora-merged-bf16.
Training data and merge summary:
- Training rows: 1650
- Train rows: 1482
- Validation rows: 168
- Navigator-full rows: 1500
- Focused-repair rows: 150
- Full corpus SHA-256:
ef7a7c9a6a99927ba72ce244e03a9da3ab86d3cf5dc70786703fb5f8bdf2a289 - Train split SHA-256:
f869d79da9ef670bc6479f8321e51b1f48cb5a16423265f34893a08e7648676e - Validation split SHA-256:
3ff7668b8216d6fa0be770d6d9ed5f1a0b12965f9312d5210b510807538738d3 - Merge method:
peft.merge_and_unload(safe_merge=True) - Merged dtype: BF16
- Base model:
nvidia/NVIDIA-Nemotron-3-Nano-4B-BF16 - Merged manifest SHA-256:
6678c0ec3a28817dba22eb9e7c682b9961f04bbfc688d0f1bcd137afaf8c8c38 - HF shard 1 SHA-256:
1d95889e945363adcd70a0be54bc29407d49e28bf7a2c0415e1732d81d64186c - HF shard 2 SHA-256:
2a2e27563e78981c130349feece291c976cf7d5384690c6327795eef6d08d4c0 - GGUF SHA-256:
7e11f2295b101e9312f97075b8e48cabd8cc89539e92c8fa4218c4973aa31d8d
The v4 full field-holdout evaluation run was local_4b_finetuned_v4_field_holdout_20260611T011930Z. A separate 50-case evidence run was local_4b_finetuned_v4_evidence_20260611T0010Z.
V5 Checkpoint
The v5 artifact was trained as figment_sft_v5 and merged from Modal checkpoint /checkpoints/figment_sft_v5/figment-sft-v5-lora into /checkpoints/figment_sft_v5/figment-sft-v5-lora-merged-bf16.
Training data and merge summary:
- Training rows: 1300
- Train rows: 1170
- Validation rows: 130
- Navigator-full rows: 1100
- Focused-repair rows: 200
- Full corpus SHA-256:
3abc2dcb1f972ee6f536c273de69f72abe9a42e402a3548c451e442a3fcd4535 - Train split SHA-256:
08ad6b76e958249b50bece528e0b26f5d3ef090166d7e5e0d48ddc46101496c7 - Validation split SHA-256:
54aadd55ab41f00880483ff0beb08c9602aae23933efcabd328d1769617fbc1a - Merge method:
peft.merge_and_unload(safe_merge=True) - Merged dtype: BF16
- Base model:
nvidia/NVIDIA-Nemotron-3-Nano-4B-BF16 - GGUF LFS SHA-256:
c7f9b38d267c2ab2b791b613e0227ce3d057e61b57b568b16ca501f2e516379c
The v5 field-holdout run was figment_sft_v5_field_workflow_holdout_modal_gpu_20260611_h100_gguf; it is retained as a regression artifact because it scored only 2/150 competence successes despite passing final JSON validation.
V6 Checkpoint
The v6 artifact was trained as figment_sft_v6 and merged from Modal checkpoint /checkpoints/figment_sft_v6/figment-sft-v6-lora into /checkpoints/figment_sft_v6/figment-sft-v6-lora-merged-bf16.
Training data and merge summary:
- Training rows: 2000
- Train rows: 1800
- Validation rows: 200
- Navigator-full rows: 1180
- Focused-repair rows: 820
- Full corpus SHA-256:
268cb36d0d36697006609f346b76c79dbf127f82837f5a1f76d47059b031c595 - Train split SHA-256:
b750779104e80a8a92c86437f9515da7a4ab97bc866c1e87f4d95fca269ab9c2 - Validation split SHA-256:
ca388117f77325a57c70af7d69145b429bd443a5ae134ce1ab419373154e25cf - Merge method:
peft.merge_and_unload(safe_merge=True) - Merged dtype: BF16
- Base model:
nvidia/NVIDIA-Nemotron-3-Nano-4B-BF16 - GGUF LFS SHA-256:
92fb2bb4a8686230f050c1696e6df749fe49ec4d41221ab9100785afa7e34009
The v6 field-holdout run was figment_sft_v6_field_workflow_holdout_modal_gpu_20260611_h100_gguf.
V7 Checkpoint
The v7 artifact was trained as figment_sft_v7 and merged from Modal checkpoint /checkpoints/figment_sft_v7/figment-sft-v7-lora into /checkpoints/figment_sft_v7/figment-sft-v7-lora-merged-bf16.
Training data and merge summary:
- Training rows: 2800
- Train rows: 2520
- Validation rows: 280
- Navigator-full rows: 1740
- Focused-repair rows: 1060
- Full corpus SHA-256:
b8bc3830beb38577047dbb2b9760aa2845234e25f41457fbfc5ce25bb6821ac0 - Train split SHA-256:
283615b21446346a9090ad6d45e750f5812222625ddaa5d2a83a15f663cb7d04 - Validation split SHA-256:
fe7b683f5007ff1f3eaac2632c9d407a8671d23c944b17b192eae964c0bbaa8d - Merge method:
peft.merge_and_unload(safe_merge=True) - Merged dtype: BF16
- Base model:
nvidia/NVIDIA-Nemotron-3-Nano-4B-BF16 - GGUF LFS SHA-256:
d85f9dd7137453035ae8ec96bcee1998358ad5975bb9c842fe9b7a077c4002b9
The v7 field-holdout run was figment_sft_v7_field_workflow_holdout_modal_gpu_20260612_h100_gguf.
V8-V14p Checkpoints
The v8-v14p artifacts continue the corrected field-workflow training loop. Each checkpoint was merged from its Modal LoRA adapter into the same BF16 base with peft.merge_and_unload(safe_merge=True) and converted to BF16 GGUF for local llama.cpp evaluation.
Training Data
The model artifacts use synthetic and de-identified datasets generated inside the Figment project. Published training corpora are available in the dataset repository build-small-hackathon/figment-eval-traces under configs figment_sft_v1 through figment_sft_v14p. The dataset files are not duplicated in this model repository.
The examples were synthetic. They were designed to teach Figment's harness behavior, not to store medical knowledge. They included full navigator outputs and focused repair tasks for schema, citations/pathways, SBAR handoff fields, missing observations, protocol urgency, and forbidden clinical language.
Evaluation
For later eval-trace artifacts, see the dataset repository build-small-hackathon/figment-eval-traces.
Observed v2 locked-harness evaluation:
Observed v3 field-holdout evaluation:
Observed v4 evaluations:
Observed v5-v7 field-holdout evaluations:
Observed v8-v14p corrected field-holdout evaluations:
Safety and Limitations
- Prototype only; not a medical device.
- Synthetic/de-identified scenarios only.
- The model must not diagnose, prescribe, dose medication, or autonomously triage.
- Deterministic red-flag rules and validators remain part of the Figment runtime. The model artifact alone is not the full safety system.
- Outputs require trained responder review and local protocol/supervisor/clinician judgment.
- The checkpoints may produce malformed, incomplete, unsupported, or overconfident outputs without the Figment harness.
License and Attribution
This archive is derived from nvidia/NVIDIA-Nemotron-3-Nano-4B-BF16 and is governed by the same upstream NVIDIA Nemotron Open Model License. Review the upstream model card and license before reuse. The Figment application code is Apache-2.0, and Figment synthetic datasets are documented separately as CC-BY-4.0 where published.
Citation
No paper is associated with these artifacts. Please cite the base model according to NVIDIA's guidance and cite this repository if using the Figment artifacts directly.
