Add model and tokenizer

Browse files

Files changed (6) hide show

README.md +49 -1
config.json +37 -0
pytorch_model.bin +3 -0
special_tokens_map.json +9 -0
tokenizer.json +0 -0
tokenizer_config.json +5 -0

README.md CHANGED Viewed

	@@ -1 +1,49 @@
1	- ~~# Megatron-BERT-large Swedish 165k for zero-shot classification~~

+---
+pipeline_tag: zero-shot-classification
+tags:
+  - zero-shot-classification
+  - swedish
+  - megatron-bert
+language:
+  - sv
+datasets:
+  - KBLab/overlim
+widget:
+  - example_title: Zero-shot
+    text: Många skjuter upp sina tandläkarbesök
+    candidate_labels: hälsa, politik, sport, religion
+inference:
+  parameters:
+    hypothesis_template: Detta exempel handlar om {}.
+---
+# Megatron-BERT-large Swedish 165k for zero-shot classification
+This model is based on Megatron-BERT-large-165k](https://huggingface.co/KBLab/megatron-bert-large-swedish-cased-165). It was fine-tuned on the QNLI task and further fine-tuned on the MNLI task.
+The model can be used with the Hugging Face zero-shot classification pipeline.
+## Usage
+```python
+>>> from transformers import pipeline
+>>> classifier = pipeline(
+...     "zero-shot-classification",
+...     model="KBlab/megatron-bert-large-swedish-cased-165k-zero-shot"
+... )
+>>> classifier(
+...     "Ruben Östlunds ”Triangle of sadness” nomineras till en Golden Globe i kategorin bästa musikal eller komedi.",
+...     candidate_labels=["hälsa", "politik", "sport", "religion", "nöje"],
+...     hypothesis_template="Detta exempel handlar om {}.",
+... )
+{'sequence': 'Ruben Östlunds ”Triangle of sadness” nomineras till en Golden Globe i kategorin bästa musikal eller komedi.',
+ 'labels': ['nöje', 'sport', 'religion', 'hälsa', 'politik'],
+ 'scores': [0.9274595379829407,
+  0.025105971843004227,
+  0.018440095707774162,
+  0.017049923539161682,
+  0.011944468133151531]}
+```

config.json ADDED Viewed

	@@ -0,0 +1,37 @@

+{
+  "_name_or_path": "megatron-bert-large-swedish-cased-165k.QM",
+  "architectures": [
+    "MegatronBertForSequenceClassification"
+  ],
+  "attention_probs_dropout_prob": 0.1,
+  "finetuning_task": "mnli",
+  "hidden_act": "gelu",
+  "hidden_dropout_prob": 0.1,
+  "hidden_size": 1024,
+  "id2label": {
+    "0": "entailment",
+    "1": "neutral",
+    "2": "contradiction"
+  },
+  "initializer_range": 0.02,
+  "intermediate_size": 4096,
+  "label2id": {
+    "contradiction": 2,
+    "entailment": 0,
+    "neutral": 1
+  },
+  "layer_norm_eps": 1e-12,
+  "max_position_embeddings": 512,
+  "model_type": "megatron-bert",
+  "num_attention_heads": 16,
+  "num_hidden_layers": 24,
+  "pad_token_id": 0,
+  "position_embedding_type": "absolute",
+  "problem_type": "single_label_classification",
+  "tokenizer_type": "BertWordPieceCase",
+  "torch_dtype": "float32",
+  "transformers_version": "4.24.0",
+  "type_vocab_size": 2,
+  "use_cache": true,
+  "vocab_size": 64128
+}

pytorch_model.bin ADDED Viewed

	@@ -0,0 +1,3 @@

+version https://git-lfs.github.com/spec/v1
+oid sha256:44b1af75b8f22314e42c7d49d4b047691c0cc0827ac44e304c495aa48dd9e596
+size 1478367541

special_tokens_map.json ADDED Viewed

	@@ -0,0 +1,9 @@

+{
+  "bos_token": "[CLS]",
+  "cls_token": "[CLS]",
+  "eos_token": "[SEP]",
+  "mask_token": "[MASK]",
+  "pad_token": "[PAD]",
+  "sep_token": "[SEP]",
+  "unk_token": "[UNK]"
+}

tokenizer.json ADDED Viewed

The diff for this file is too large to render. See raw diff

tokenizer_config.json ADDED Viewed

	@@ -0,0 +1,5 @@

+{
+  "name_or_path": "megatron-bert-large-swedish-cased-165k.QM",
+  "special_tokens_map_file": "pretrained_model_hf_large_165K/special_tokens_map.json",
+  "tokenizer_class": "PreTrainedTokenizerFast"
+}