CyrexPro commited on
Commit
5be2f6c
1 Parent(s): 75503dd

Model save

Browse files
README.md ADDED
@@ -0,0 +1,71 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ ---
2
+ base_model: google/pegasus-xsum
3
+ tags:
4
+ - generated_from_trainer
5
+ metrics:
6
+ - rouge
7
+ model-index:
8
+ - name: pegasus-xsum-finetuned-cnn_dailymail
9
+ results: []
10
+ ---
11
+
12
+ <!-- This model card has been generated automatically according to the information the Trainer had access to. You
13
+ should probably proofread and complete it, then remove this comment. -->
14
+
15
+ # pegasus-xsum-finetuned-cnn_dailymail
16
+
17
+ This model is a fine-tuned version of [google/pegasus-xsum](https://huggingface.co/google/pegasus-xsum) on the None dataset.
18
+ It achieves the following results on the evaluation set:
19
+ - Loss: 0.8958
20
+ - Rouge1: 45.7795
21
+ - Rouge2: 23.3182
22
+ - Rougel: 32.9241
23
+ - Rougelsum: 42.3126
24
+ - Bleu 1: 35.4715
25
+ - Bleu 2: 24.0726
26
+ - Bleu 3: 17.9591
27
+ - Meteor: 32.8897
28
+ - Lungime rezumat: 43.3773
29
+ - Lungime original: 48.6937
30
+
31
+ ## Model description
32
+
33
+ More information needed
34
+
35
+ ## Intended uses & limitations
36
+
37
+ More information needed
38
+
39
+ ## Training and evaluation data
40
+
41
+ More information needed
42
+
43
+ ## Training procedure
44
+
45
+ ### Training hyperparameters
46
+
47
+ The following hyperparameters were used during training:
48
+ - learning_rate: 5.6e-05
49
+ - train_batch_size: 4
50
+ - eval_batch_size: 4
51
+ - seed: 42
52
+ - optimizer: Adam with betas=(0.9,0.999) and epsilon=1e-08
53
+ - lr_scheduler_type: linear
54
+ - num_epochs: 4
55
+
56
+ ### Training results
57
+
58
+ | Training Loss | Epoch | Step | Validation Loss | Rouge1 | Rouge2 | Rougel | Rougelsum | Bleu 1 | Bleu 2 | Bleu 3 | Meteor | Lungime rezumat | Lungime original |
59
+ |:-------------:|:-----:|:-----:|:---------------:|:-------:|:-------:|:-------:|:---------:|:-------:|:-------:|:-------:|:-------:|:---------------:|:----------------:|
60
+ | 1.1281 | 1.0 | 14330 | 0.9373 | 44.64 | 22.2111 | 32.0228 | 41.1223 | 34.4946 | 23.079 | 17.0673 | 31.8685 | 43.543 | 48.6937 |
61
+ | 0.9091 | 2.0 | 28660 | 0.9095 | 45.0713 | 22.7428 | 32.4247 | 41.554 | 34.9397 | 23.5631 | 17.5094 | 32.1814 | 43.3467 | 48.6937 |
62
+ | 0.8455 | 3.0 | 42990 | 0.8982 | 45.5457 | 23.1315 | 32.7153 | 42.0349 | 35.2659 | 23.8773 | 17.8174 | 32.7185 | 43.5743 | 48.6937 |
63
+ | 0.8076 | 4.0 | 57320 | 0.8958 | 45.7795 | 23.3182 | 32.9241 | 42.3126 | 35.4715 | 24.0726 | 17.9591 | 32.8897 | 43.3773 | 48.6937 |
64
+
65
+
66
+ ### Framework versions
67
+
68
+ - Transformers 4.40.0
69
+ - Pytorch 2.2.2+cu118
70
+ - Datasets 2.19.0
71
+ - Tokenizers 0.19.1
generation_config.json ADDED
@@ -0,0 +1,11 @@
 
 
 
 
 
 
 
 
 
 
 
 
1
+ {
2
+ "bos_token_id": 0,
3
+ "decoder_start_token_id": 0,
4
+ "eos_token_id": 1,
5
+ "forced_eos_token_id": 1,
6
+ "length_penalty": 0.6,
7
+ "max_length": 64,
8
+ "num_beams": 8,
9
+ "pad_token_id": 0,
10
+ "transformers_version": "4.40.0"
11
+ }
model.safetensors CHANGED
@@ -1,3 +1,3 @@
1
  version https://git-lfs.github.com/spec/v1
2
- oid sha256:070a09f7011ef19da0a9127cdaea0a18e0669b63d931d260cfad1545f0916619
3
  size 2279458540
 
1
  version https://git-lfs.github.com/spec/v1
2
+ oid sha256:0620c40c24c77afec355474f60fdf15cd2a0df00c4f1ebd0887476ea7e61c641
3
  size 2279458540
runs/May08_01-12-29_DESKTOP-I570M0U/events.out.tfevents.1715119951.DESKTOP-I570M0U.76844.0 CHANGED
@@ -1,3 +1,3 @@
1
  version https://git-lfs.github.com/spec/v1
2
- oid sha256:ed07ffb7cd32502ae8968f252d5df5bd3583e21f80e5721f003047f0b2c23f24
3
- size 8698
 
1
  version https://git-lfs.github.com/spec/v1
2
+ oid sha256:235969b37d9a2f885d941e1cadb2906d135a90bd6b0b82647de5fe6da4faedc2
3
+ size 10081