Improved the hyperparameters
Browse files- README.md +15 -11
- model.safetensors +1 -1
README.md
CHANGED
@@ -17,12 +17,12 @@ should probably proofread and complete it, then remove this comment. -->
|
|
17 |
|
18 |
This model is a fine-tuned version of [google-t5/t5-small](https://huggingface.co/google-t5/t5-small) on an unknown dataset.
|
19 |
It achieves the following results on the evaluation set:
|
20 |
-
- Loss: 0.
|
21 |
-
- Rouge1: 0.
|
22 |
-
- Rouge2: 0.
|
23 |
-
- Rougel: 0.
|
24 |
-
- Rougelsum: 0.
|
25 |
-
- Gen Len: 6.
|
26 |
|
27 |
## Model description
|
28 |
|
@@ -41,21 +41,25 @@ More information needed
|
|
41 |
### Training hyperparameters
|
42 |
|
43 |
The following hyperparameters were used during training:
|
44 |
-
- learning_rate:
|
45 |
- train_batch_size: 16
|
46 |
- eval_batch_size: 16
|
47 |
- seed: 42
|
48 |
- optimizer: Adam with betas=(0.9,0.999) and epsilon=1e-08
|
49 |
- lr_scheduler_type: linear
|
50 |
-
-
|
|
|
51 |
|
52 |
### Training results
|
53 |
|
54 |
| Training Loss | Epoch | Step | Validation Loss | Rouge1 | Rouge2 | Rougel | Rougelsum | Gen Len |
|
55 |
|:-------------:|:-------:|:----:|:---------------:|:------:|:------:|:------:|:---------:|:-------:|
|
56 |
-
|
|
57 |
-
|
|
58 |
-
| 0.
|
|
|
|
|
|
|
59 |
|
60 |
|
61 |
### Framework versions
|
|
|
17 |
|
18 |
This model is a fine-tuned version of [google-t5/t5-small](https://huggingface.co/google-t5/t5-small) on an unknown dataset.
|
19 |
It achieves the following results on the evaluation set:
|
20 |
+
- Loss: 0.9997
|
21 |
+
- Rouge1: 0.7404
|
22 |
+
- Rouge2: 0.6249
|
23 |
+
- Rougel: 0.7403
|
24 |
+
- Rougelsum: 0.7413
|
25 |
+
- Gen Len: 6.9017
|
26 |
|
27 |
## Model description
|
28 |
|
|
|
41 |
### Training hyperparameters
|
42 |
|
43 |
The following hyperparameters were used during training:
|
44 |
+
- learning_rate: 5e-05
|
45 |
- train_batch_size: 16
|
46 |
- eval_batch_size: 16
|
47 |
- seed: 42
|
48 |
- optimizer: Adam with betas=(0.9,0.999) and epsilon=1e-08
|
49 |
- lr_scheduler_type: linear
|
50 |
+
- lr_scheduler_warmup_steps: 1000
|
51 |
+
- num_epochs: 32
|
52 |
|
53 |
### Training results
|
54 |
|
55 |
| Training Loss | Epoch | Step | Validation Loss | Rouge1 | Rouge2 | Rougel | Rougelsum | Gen Len |
|
56 |
|:-------------:|:-------:|:----:|:---------------:|:------:|:------:|:------:|:---------:|:-------:|
|
57 |
+
| 2.3461 | 3.7879 | 500 | 1.0711 | 0.7407 | 0.6192 | 0.736 | 0.7374 | 7.188 |
|
58 |
+
| 1.0075 | 7.5758 | 1000 | 0.9645 | 0.7313 | 0.6071 | 0.7304 | 0.7303 | 6.9274 |
|
59 |
+
| 0.7921 | 11.3636 | 1500 | 0.9563 | 0.7306 | 0.6079 | 0.7323 | 0.7325 | 6.7863 |
|
60 |
+
| 0.6587 | 15.1515 | 2000 | 0.9697 | 0.7382 | 0.6142 | 0.739 | 0.7397 | 6.8675 |
|
61 |
+
| 0.5579 | 18.9394 | 2500 | 0.9905 | 0.7388 | 0.6203 | 0.7378 | 0.7395 | 6.8718 |
|
62 |
+
| 0.4984 | 22.7273 | 3000 | 0.9997 | 0.7404 | 0.6249 | 0.7403 | 0.7413 | 6.9017 |
|
63 |
|
64 |
|
65 |
### Framework versions
|
model.safetensors
CHANGED
@@ -1,3 +1,3 @@
|
|
1 |
version https://git-lfs.github.com/spec/v1
|
2 |
-
oid sha256:
|
3 |
size 241988648
|
|
|
1 |
version https://git-lfs.github.com/spec/v1
|
2 |
+
oid sha256:a929e79f3a3ef1f8d357a26daa1ef2784918613914d20edadf972ac66d3f41cc
|
3 |
size 241988648
|