Transformers
Polish
tokenizer
fast-tokenizer
polish
Inference Endpoints
Edit model card

This is polish fast tokenizer.

Number of documents used to train tokenizer:

  • 25 088 398

Sample usge with transformers:

from transformers import AutoTokenizer

tokenizer = AutoTokenizer.from_pretrained('radlab/polish-fast-tokenizer')
tokenizer.decode(tokenizer("Ala ma kota i psa").input_ids)
Downloads last month
0
Unable to determine this model’s pipeline type. Check the docs .

Datasets used to train radlab/polish-fast-tokenizer