NexaAIDev/Octopus-v2 · Function Scalability?

It is really great to see certain work which really focus on improving life efficiency without acting as a chatbot.

It seems you are tokenizing the functions (you think this is very important) and train only on those functions to let model learn how to use these tools by model parameters rather than paste them in prompt. This do save the VRAM since it is working in backend.

But I want to know whether you have some "best practice" to add more functions later?
And why you think making it a special token in the tokenizer (256022) is better than just a short special free text while keep the original tokenizer (256000)?