Hugging Face inference endpoints

#18

by AlexNevo - opened Jul 17, 2024

Discussion

AlexNevo

Jul 17, 2024

•

edited Jul 17, 2024

Hi,
I would like to configure a Hugging Face inference endpoint to deploy this model. What would you recommend ? Expecially for the section about Max Input Length (per Query), Max Number of Tokens (per Query), Max Batch Prefill Tokens and Max Batch Total Tokens considering that my DDL is about 4500 tokens long :

wongjingping

Defog.ai org Jul 18, 2024

Hi @AlexNevo , you may refer to their documentation for how to set the parameters.
I believe you would need to increase the max input length given your large DDL size.
https://huggingface.co/docs/inference-endpoints/index

Upload images, audio, and videos by dragging in the text input, pasting, or clicking here.

Tap or paste here to upload images

· Sign up or log in to comment