kxx-kkk commited on
Commit
6ffeacb
verified
1 Parent(s): a9d929a

End of training

Browse files
README.md ADDED
@@ -0,0 +1,66 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ ---
2
+ license: cc-by-4.0
3
+ base_model: kxx-kkk/FYP_sq2_mrqa_adqa_synqa
4
+ tags:
5
+ - generated_from_trainer
6
+ model-index:
7
+ - name: FYP_sq2_mrqa_adqa_synqa_quoref
8
+ results: []
9
+ ---
10
+
11
+ <!-- This model card has been generated automatically according to the information the Trainer had access to. You
12
+ should probably proofread and complete it, then remove this comment. -->
13
+
14
+ # FYP_sq2_mrqa_adqa_synqa_quoref
15
+
16
+ This model is a fine-tuned version of [kxx-kkk/FYP_sq2_mrqa_adqa_synqa](https://huggingface.co/kxx-kkk/FYP_sq2_mrqa_adqa_synqa) on an unknown dataset.
17
+ It achieves the following results on the evaluation set:
18
+ - Loss: 1.8654
19
+
20
+ ## Model description
21
+
22
+ More information needed
23
+
24
+ ## Intended uses & limitations
25
+
26
+ More information needed
27
+
28
+ ## Training and evaluation data
29
+
30
+ More information needed
31
+
32
+ ## Training procedure
33
+
34
+ ### Training hyperparameters
35
+
36
+ The following hyperparameters were used during training:
37
+ - learning_rate: 1e-05
38
+ - train_batch_size: 16
39
+ - eval_batch_size: 16
40
+ - seed: 42
41
+ - optimizer: Adam with betas=(0.9,0.999) and epsilon=1e-08
42
+ - lr_scheduler_type: linear
43
+ - num_epochs: 3
44
+ - mixed_precision_training: Native AMP
45
+
46
+ ### Training results
47
+
48
+ | Training Loss | Epoch | Step | Validation Loss |
49
+ |:-------------:|:-----:|:----:|:---------------:|
50
+ | 2.3662 | 0.32 | 500 | 2.2585 |
51
+ | 2.1791 | 0.64 | 1000 | 2.1171 |
52
+ | 2.1006 | 0.95 | 1500 | 2.0172 |
53
+ | 2.0647 | 1.27 | 2000 | 1.9514 |
54
+ | 2.0247 | 1.59 | 2500 | 1.9741 |
55
+ | 1.9766 | 1.91 | 3000 | 1.8697 |
56
+ | 1.9395 | 2.23 | 3500 | 1.8859 |
57
+ | 1.9254 | 2.54 | 4000 | 1.8358 |
58
+ | 1.9118 | 2.86 | 4500 | 1.8654 |
59
+
60
+
61
+ ### Framework versions
62
+
63
+ - Transformers 4.39.3
64
+ - Pytorch 2.2.1+cu121
65
+ - Datasets 2.18.0
66
+ - Tokenizers 0.15.2
model.safetensors CHANGED
@@ -1,3 +1,3 @@
1
  version https://git-lfs.github.com/spec/v1
2
- oid sha256:9f467d1f7087dff7c25aa9d9d644c39e29ee1cbe1302205abc42901f78705f42
3
  size 735356752
 
1
  version https://git-lfs.github.com/spec/v1
2
+ oid sha256:3f30623ffef7896415db95138867595e40e35ae15409a4e9828cebd7ac3a6ed1
3
  size 735356752
runs/Apr07_12-50-26_b850becae0cc/events.out.tfevents.1712494242.b850becae0cc.2399.0 CHANGED
@@ -1,3 +1,3 @@
1
  version https://git-lfs.github.com/spec/v1
2
- oid sha256:3c5aebd7617e5bd702c5aca221a4c96952445cd0df8417eeae5bb26d3940b21e
3
- size 9399
 
1
  version https://git-lfs.github.com/spec/v1
2
+ oid sha256:e2329a215116661ed053f2b9b306e0209d67f88e0216153248b20cae2f2b6832
3
+ size 9753