gpt2-baseline / tokenizer_config.json
Shaojie Jiang
GPT2 small finetuned on Wikitext-103, as a baseline for CT.
73da027
{"unk_token": "<|endoftext|>", "bos_token": "<|endoftext|>", "eos_token": "<|endoftext|>", "add_prefix_space": false, "model_max_length": 1024, "special_tokens_map_file": null, "name_or_path": "gpt2", "tokenizer_class": "GPT2Tokenizer"}