Any-to-Any
Transformers
Safetensors
gemma4_assistant
text-generation
fp8
vllm
llm-compressor
compressed-tensors
speculative-decoding
mtp
gemma4
Instructions to use BarraHome/gemma-4-31B-it-assistant-FP8-dynamic with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use BarraHome/gemma-4-31B-it-assistant-FP8-dynamic with Transformers:
# Load model directly from transformers import AutoTokenizer, AutoModelForCausalLM tokenizer = AutoTokenizer.from_pretrained("BarraHome/gemma-4-31B-it-assistant-FP8-dynamic") model = AutoModelForCausalLM.from_pretrained("BarraHome/gemma-4-31B-it-assistant-FP8-dynamic", device_map="auto") - Notebooks
- Google Colab
- Kaggle
Welcome to the community
The community tab is the place to discuss and collaborate with the HF community!