Team Ai
Modelpublic

kasperius/falcon-7b-sharded-bf16-finetuned-html-code-generation

sourceHugging Faceupdated 2y agoView on Hugging Face
1likes18downloads
Model Card

<!-- This model card has been generated automatically according to the information the Trainer had access to. You should probably proofread and complete it, then remove this comment. -->

falcon-7b-sharded-bf16-finetuned-html-code-generation

This model is a fine-tuned version of ybelkada/falcon-7b-sharded-bf16 on an unknown dataset. It achieves the following results on the evaluation set:

  • —Loss: 0.7322

Model description

More information needed

Intended uses & limitations

More information needed

Training and evaluation data

More information needed

Training procedure

Training hyperparameters

The following hyperparameters were used during training:

  • —learning_rate: 0.0002
  • —trainbatchsize: 2
  • —evalbatchsize: 8
  • —seed: 42
  • —gradientaccumulationsteps: 2
  • —totaltrainbatch_size: 4
  • —optimizer: Adam with betas=(0.9,0.999) and epsilon=1e-08
  • —lrschedulertype: cosine
  • —lrschedulerwarmup_ratio: 0.03
  • —training_steps: 320

Training results

Training LossEpochStepValidation Loss
No log0.1794201.8071
No log0.3587401.4823
No log0.5381601.3637
No log0.7175801.2700
No log0.89691001.2054
No log1.07621201.1352
No log1.25561401.1297
No log1.43501601.0126
No log1.61431800.9738
No log1.79372000.9058
No log1.97312200.8581
No log2.15252400.7948
No log2.33182600.7601
No log2.51122800.7397
No log2.69063000.7332
No log2.87003200.7322

Framework versions

  • —PEFT 0.12.1.dev0
  • —Transformers 4.43.3
  • —Pytorch 2.3.1+cu121
  • —Datasets 2.20.0
  • —Tokenizers 0.19.1