josueu commited on
Commit
657cc1d
·
verified ·
1 Parent(s): 919b083

End of training

Browse files
Files changed (2) hide show
  1. README.md +76 -0
  2. model.safetensors +1 -1
README.md ADDED
@@ -0,0 +1,76 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ ---
2
+ library_name: transformers
3
+ license: apache-2.0
4
+ base_model: Helsinki-NLP/opus-mt-es-en
5
+ tags:
6
+ - generated_from_trainer
7
+ metrics:
8
+ - bleu
9
+ model-index:
10
+ - name: MarianMT-finetuned-ES-ZAP-v4
11
+ results: []
12
+ ---
13
+
14
+ <!-- This model card has been generated automatically according to the information the Trainer had access to. You
15
+ should probably proofread and complete it, then remove this comment. -->
16
+
17
+ # MarianMT-finetuned-ES-ZAP-v4
18
+
19
+ This model is a fine-tuned version of [Helsinki-NLP/opus-mt-es-en](https://huggingface.co/Helsinki-NLP/opus-mt-es-en) on the None dataset.
20
+ It achieves the following results on the evaluation set:
21
+ - Loss: 0.2469
22
+ - Bleu: 20.5832
23
+ - Meteor: 0.4548
24
+ - Ter: 64.5992
25
+ - Chrf: 43.7279
26
+
27
+ ## Model description
28
+
29
+ More information needed
30
+
31
+ ## Intended uses & limitations
32
+
33
+ More information needed
34
+
35
+ ## Training and evaluation data
36
+
37
+ More information needed
38
+
39
+ ## Training procedure
40
+
41
+ ### Training hyperparameters
42
+
43
+ The following hyperparameters were used during training:
44
+ - learning_rate: 1e-05
45
+ - train_batch_size: 8
46
+ - eval_batch_size: 8
47
+ - seed: 42
48
+ - optimizer: Use OptimizerNames.ADAMW_TORCH_FUSED with betas=(0.9,0.999) and epsilon=1e-08 and optimizer_args=No additional optimizer arguments
49
+ - lr_scheduler_type: linear
50
+ - num_epochs: 50
51
+
52
+ ### Training results
53
+
54
+ | Training Loss | Epoch | Step | Validation Loss | Bleu | Meteor | Ter | Chrf |
55
+ |:-------------:|:-----:|:----:|:---------------:|:-------:|:------:|:--------:|:-------:|
56
+ | 1.3736 | 1.0 | 179 | 0.5978 | 0.3111 | 0.0882 | 216.7939 | 9.2974 |
57
+ | 0.4949 | 2.0 | 358 | 0.4382 | 3.6416 | 0.1889 | 95.1336 | 20.9865 |
58
+ | 0.3762 | 3.0 | 537 | 0.3719 | 6.2974 | 0.2275 | 88.7405 | 24.9969 |
59
+ | 0.3149 | 4.0 | 716 | 0.3382 | 8.2395 | 0.2822 | 84.8282 | 29.5532 |
60
+ | 0.2765 | 5.0 | 895 | 0.3156 | 8.4096 | 0.3200 | 79.4847 | 32.3501 |
61
+ | 0.2462 | 6.0 | 1074 | 0.2968 | 11.4876 | 0.3457 | 75.8588 | 35.1628 |
62
+ | 0.2217 | 7.0 | 1253 | 0.2841 | 13.5969 | 0.3650 | 77.0992 | 36.6277 |
63
+ | 0.2013 | 8.0 | 1432 | 0.2731 | 15.4573 | 0.3967 | 71.3740 | 38.5145 |
64
+ | 0.1842 | 9.0 | 1611 | 0.2638 | 16.6958 | 0.4135 | 70.1336 | 40.9340 |
65
+ | 0.1690 | 10.0 | 1790 | 0.2573 | 17.4951 | 0.4206 | 67.4618 | 41.4994 |
66
+ | 0.1560 | 11.0 | 1969 | 0.2531 | 18.0435 | 0.4313 | 67.6527 | 41.9272 |
67
+ | 0.1446 | 12.0 | 2148 | 0.2469 | 20.5832 | 0.4548 | 64.5992 | 43.7279 |
68
+ | 0.1331 | 13.0 | 2327 | 0.2449 | 18.6926 | 0.4423 | 66.6985 | 43.3962 |
69
+
70
+
71
+ ### Framework versions
72
+
73
+ - Transformers 5.0.0
74
+ - Pytorch 2.10.0+cu128
75
+ - Datasets 4.0.0
76
+ - Tokenizers 0.22.2
model.safetensors CHANGED
@@ -1,3 +1,3 @@
1
  version https://git-lfs.github.com/spec/v1
2
- oid sha256:a0dc378feb6bf6b32cdfee0c6153695b73acecb11b0e0b53b4d537020e62fad9
3
  size 576209740
 
1
  version https://git-lfs.github.com/spec/v1
2
+ oid sha256:77a0def1678c88589622beef898387ffecbabeb8d41d8ab787a08cac55ec8bdd
3
  size 576209740