The purpose of this model is to help experiment BPE tokenization with a very large vocabulary. Applications include: text compression, language models, search engines. I recommend gigatoken to run this model: https://github.com/marcelroed/gigatoken .
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support