Dev-SriramB/DPB_buster

Browse files

Files changed (3) hide show

README.md +11 -19
adapter_model.safetensors +1 -1
training_args.bin +2 -2

README.md CHANGED Viewed

@@ -1,9 +1,9 @@
 ---
-license: apache-2.0
 library_name: peft
 tags:
 - generated_from_trainer
-base_model: TheBloke/Mistral-7B-Instruct-v0.2-GPTQ
 model-index:
 - name: balagpt-ft2
   results: []
@@ -16,7 +16,7 @@ should probably proofread and complete it, then remove this comment. -->
 This model is a fine-tuned version of [TheBloke/Mistral-7B-Instruct-v0.2-GPTQ](https://huggingface.co/TheBloke/Mistral-7B-Instruct-v0.2-GPTQ) on the None dataset.
 It achieves the following results on the evaluation set:
-- Loss: 0.6050
 ## Model description
@@ -44,29 +44,21 @@ The following hyperparameters were used during training:
 - optimizer: Adam with betas=(0.9,0.999) and epsilon=1e-08
 - lr_scheduler_type: linear
 - lr_scheduler_warmup_steps: 2
-- num_epochs: 10
 - mixed_precision_training: Native AMP
 ### Training results
 | Training Loss | Epoch | Step | Validation Loss |
 |:-------------:|:-----:|:----:|:---------------:|
-| 2.3314        | 1.0   | 5    | 1.9312          |
-| 1.7381        | 2.0   | 10   | 1.4899          |
-| 1.3108        | 3.0   | 15   | 1.1581          |
-| 0.9634        | 4.0   | 20   | 0.9015          |
-| 0.718         | 5.0   | 25   | 0.7503          |
-| 0.5738        | 6.0   | 30   | 0.6664          |
-| 0.4784        | 7.0   | 35   | 0.6163          |
-| 0.4266        | 8.0   | 40   | 0.6102          |
-| 0.4016        | 9.0   | 45   | 0.6052          |
-| 0.3864        | 10.0  | 50   | 0.6050          |
 ### Framework versions
-- PEFT 0.10.0
-- Transformers 4.39.3
-- Pytorch 2.1.0+cu121
-- Datasets 2.19.0
-- Tokenizers 0.15.2

 ---
+base_model: TheBloke/Mistral-7B-Instruct-v0.2-GPTQ
 library_name: peft
+license: apache-2.0
 tags:
 - generated_from_trainer
 model-index:
 - name: balagpt-ft2
   results: []
 This model is a fine-tuned version of [TheBloke/Mistral-7B-Instruct-v0.2-GPTQ](https://huggingface.co/TheBloke/Mistral-7B-Instruct-v0.2-GPTQ) on the None dataset.
 It achieves the following results on the evaluation set:
+- Loss: 0.7072
 ## Model description
 - optimizer: Adam with betas=(0.9,0.999) and epsilon=1e-08
 - lr_scheduler_type: linear
 - lr_scheduler_warmup_steps: 2
+- num_epochs: 2
 - mixed_precision_training: Native AMP
 ### Training results
 | Training Loss | Epoch | Step | Validation Loss |
 |:-------------:|:-----:|:----:|:---------------:|
+| 0.5677        | 1.0   | 175  | 0.6737          |
+| 0.2266        | 2.0   | 350  | 0.7072          |
 ### Framework versions
+- PEFT 0.13.0
+- Transformers 4.44.2
+- Pytorch 2.4.1+cu121
+- Datasets 3.0.1
+- Tokenizers 0.19.1

adapter_model.safetensors CHANGED Viewed

@@ -1,3 +1,3 @@
 version https://git-lfs.github.com/spec/v1
-oid sha256:a5064f0c517c83635d202bde21aaed816bebf91bac61a6c1e509894259c547b5
 size 8397056

 version https://git-lfs.github.com/spec/v1
+oid sha256:1159a2c2cf19369defcc228607f5dd60018c868bfccb96b407aa2458afdb77db
 size 8397056

training_args.bin CHANGED Viewed

@@ -1,3 +1,3 @@
 version https://git-lfs.github.com/spec/v1
-oid sha256:e0e628017ce5b7595994d35d37ee0a2e8bb5d8bbb9cb0cfa9cf0b477bde014f4
-size 4920

 version https://git-lfs.github.com/spec/v1
+oid sha256:03547ca9e6158c53cf938382db71ead0e4656fc78145f6868a1c06a0b729289b
+size 5176