End of training

Browse files

Files changed (5) hide show

README.md +82 -0
emissions.csv +2 -0
pytorch_model.bin +1 -1
runs/Jan12_14-52-43_2cb2cefa7ce2/events.out.tfevents.1673535186.2cb2cefa7ce2.1925.0 +2 -2
runs/Jan12_14-52-43_2cb2cefa7ce2/events.out.tfevents.1673546557.2cb2cefa7ce2.1925.2 +3 -0

README.md ADDED Viewed

	@@ -0,0 +1,82 @@

+---
+license: apache-2.0
+tags:
+- generated_from_trainer
+datasets:
+- samsum
+metrics:
+- rouge
+model-index:
+- name: switch-base-16-finetuned-samsum
+  results:
+  - task:
+      name: Sequence-to-sequence Language Modeling
+      type: text2text-generation
+    dataset:
+      name: samsum
+      type: samsum
+      config: samsum
+      split: validation
+      args: samsum
+    metrics:
+    - name: Rouge1
+      type: rouge
+      value: 47.2139
+---
+<!-- This model card has been generated automatically according to the information the Trainer had access to. You
+should probably proofread and complete it, then remove this comment. -->
+# switch-base-16-finetuned-samsum
+This model is a fine-tuned version of [google/switch-base-16](https://huggingface.co/google/switch-base-16) on the samsum dataset.
+It achieves the following results on the evaluation set:
+- Loss: 1.4434
+- Rouge1: 47.2139
+- Rouge2: 23.3399
+- Rougel: 39.8364
+- Rougelsum: 43.2592
+- Gen Len: 16.9194
+## Model description
+More information needed
+## Intended uses & limitations
+More information needed
+## Training and evaluation data
+More information needed
+## Training procedure
+### Training hyperparameters
+The following hyperparameters were used during training:
+- learning_rate: 5e-05
+- train_batch_size: 4
+- eval_batch_size: 4
+- seed: 42
+- optimizer: Adam with betas=(0.9,0.999) and epsilon=1e-08
+- lr_scheduler_type: linear
+- num_epochs: 5
+### Training results
+| Training Loss | Epoch | Step  | Validation Loss | Rouge1  | Rouge2  | Rougel  | Rougelsum | Gen Len |
+|:-------------:|:-----:|:-----:|:---------------:|:-------:|:-------:|:-------:|:---------:|:-------:|
+| 1.846         | 1.0   | 3683  | 1.4857          | 45.9134 | 22.4258 | 38.9716 | 42.6169   | 17.0623 |
+| 1.5734        | 2.0   | 7366  | 1.4346          | 47.574  | 24.2967 | 40.3749 | 44.2636   | 17.3790 |
+| 1.38          | 3.0   | 11049 | 1.4277          | 47.9915 | 24.9077 | 40.658  | 44.5301   | 17.1406 |
+| 1.2388        | 4.0   | 14732 | 1.4223          | 48.3444 | 25.4061 | 41.2776 | 45.0434   | 16.9254 |
+| 1.1629        | 5.0   | 18415 | 1.4372          | 48.5991 | 25.5464 | 41.3726 | 45.0784   | 16.9890 |
+### Framework versions
+- Transformers 4.26.0.dev0
+- Pytorch 1.13.1+cu116
+- Datasets 2.7.1
+- Tokenizers 0.13.2

emissions.csv ADDED Viewed

	@@ -0,0 +1,2 @@


1	+ timestamp,experiment_id,project_name,duration,emissions,energy_consumed,country_name,country_iso_code,region,on_cloud,cloud_provider,cloud_region
2	+ 2023-01-12T17:59:02,f503872b-5629-4375-b9a9-ac79fb936088,codecarbon,11156.288381099701,0.5907260779761291,0.9903934752534171,Spain,ESP,murcia,N,,

pytorch_model.bin CHANGED Viewed

@@ -1,3 +1,3 @@
 version https://git-lfs.github.com/spec/v1
-oid sha256:199e2c06d54a3a79ae358ef87e061b1d688de34647b8510bbf56d2948804dd56
 size 4289811489

 version https://git-lfs.github.com/spec/v1
+oid sha256:4d2ac5946ef3aad20037431d462828a37a65ed4deeec498408c4a6b037ddffbf
 size 4289811489

runs/Jan12_14-52-43_2cb2cefa7ce2/events.out.tfevents.1673535186.2cb2cefa7ce2.1925.0 CHANGED Viewed

@@ -1,3 +1,3 @@
 version https://git-lfs.github.com/spec/v1
-oid sha256:1c0fd82bccbaf584fe1d0f9ec9e059e7dcf75421895a83611e7b80447d89ae00
-size 14067

 version https://git-lfs.github.com/spec/v1
+oid sha256:338cef2e8fcb2df2cfb9e0afd7e1aa12b79517c8e7c3b121a94727bdecd765b2
+size 14427

runs/Jan12_14-52-43_2cb2cefa7ce2/events.out.tfevents.1673546557.2cb2cefa7ce2.1925.2 ADDED Viewed

	@@ -0,0 +1,3 @@

+version https://git-lfs.github.com/spec/v1
+oid sha256:daf2347e9707147b22fec507bf2fa862dc66a557acde6786de0187ae1812117b
+size 575