bala3040/paper_eval_gpt

Files changed (4) hide show

README.md CHANGED Viewed

@@ -16,7 +16,7 @@ should probably proofread and complete it, then remove this comment. -->
 This model is a fine-tuned version of [TheBloke/Mistral-7B-Instruct-v0.2-GPTQ](https://huggingface.co/TheBloke/Mistral-7B-Instruct-v0.2-GPTQ) on the None dataset.
 It achieves the following results on the evaluation set:
-- Loss: 0.5294
 ## Model description
@@ -51,16 +51,16 @@ The following hyperparameters were used during training:
 | Training Loss | Epoch | Step | Validation Loss |
 |:-------------:|:-----:|:----:|:---------------:|
-| 2.1594        | 1.0   | 5    | 1.8573          |
-| 1.6557        | 2.0   | 10   | 1.3995          |
-| 1.2039        | 3.0   | 15   | 1.0328          |
-| 0.8274        | 4.0   | 20   | 0.7538          |
-| 0.5816        | 5.0   | 25   | 0.6224          |
-| 0.4699        | 6.0   | 30   | 0.5730          |
-| 0.408         | 7.0   | 35   | 0.5498          |
-| 0.3595        | 8.0   | 40   | 0.5377          |
-| 0.3235        | 9.0   | 45   | 0.5319          |
-| 0.3042        | 10.0  | 50   | 0.5294          |
 ### Framework versions

 This model is a fine-tuned version of [TheBloke/Mistral-7B-Instruct-v0.2-GPTQ](https://huggingface.co/TheBloke/Mistral-7B-Instruct-v0.2-GPTQ) on the None dataset.
 It achieves the following results on the evaluation set:
+- Loss: 0.5969
 ## Model description
 | Training Loss | Epoch | Step | Validation Loss |
 |:-------------:|:-----:|:----:|:---------------:|
+| 2.5007        | 1.0   | 5    | 2.0719          |
+| 1.8713        | 2.0   | 10   | 1.5969          |
+| 1.4159        | 3.0   | 15   | 1.2384          |
+| 1.0399        | 4.0   | 20   | 0.9500          |
+| 0.7569        | 5.0   | 25   | 0.7594          |
+| 0.59          | 6.0   | 30   | 0.6689          |
+| 0.5004        | 7.0   | 35   | 0.6295          |
+| 0.4488        | 8.0   | 40   | 0.6123          |
+| 0.4119        | 9.0   | 45   | 0.5994          |
+| 0.386         | 10.0  | 50   | 0.5969          |
 ### Framework versions

adapter_model.safetensors CHANGED Viewed

@@ -1,3 +1,3 @@
 version https://git-lfs.github.com/spec/v1
-oid sha256:10a9947d86f2c6a34bf29d90249f8ff19f420e723f742704413e3231bedf051e
 size 8397056

 version https://git-lfs.github.com/spec/v1
+oid sha256:12b5ef7e77573cd180bdfb58950cdb99dbd6e99eedcdb45fcda8eac35081fcdc
 size 8397056

runs/Apr17_16-58-05_7403b92412ff/events.out.tfevents.1713373086.7403b92412ff.196.0 ADDED Viewed

+version https://git-lfs.github.com/spec/v1
+oid sha256:64b798da8b2b2df476fce5bee89b6c0da1a4d55dff2daa5ae7106b430df0ea14
+size 10303

training_args.bin CHANGED Viewed

@@ -1,3 +1,3 @@
 version https://git-lfs.github.com/spec/v1
-oid sha256:4cf8d48594cbcef15c2a218f4b26be8b7e8f28dcacde6925f2c94fbc3d05d247
-size 4856

 version https://git-lfs.github.com/spec/v1
+oid sha256:3830df37fb9742ae81ecd72f7d2180409ce95de1c4c24a4981091be978ca3675
+size 4920