Instructions to use unsloth/Qwen3.8-27B-NVFP4 with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Local Apps Settings
- Unsloth Studio
How to use unsloth/Qwen3.8-27B-NVFP4 with Unsloth Studio:
Install Unsloth Studio (macOS, Linux, WSL)
curl -fsSL https://unsloth.ai/install.sh | sh # Run unsloth studio unsloth studio -H 0.0.0.0 -p 8888 # Then open http://localhost:8888 in your browser # Search for unsloth/Qwen3.8-27B-NVFP4 to start chatting
Install Unsloth Studio (Windows)
irm https://unsloth.ai/install.ps1 | iex # Run unsloth studio unsloth studio -H 0.0.0.0 -p 8888 # Then open http://localhost:8888 in your browser # Search for unsloth/Qwen3.8-27B-NVFP4 to start chatting
Using HuggingFace Spaces for Unsloth
# No setup required # Open https://huggingface.co/spaces/unsloth/studio in your browser # Search for unsloth/Qwen3.8-27B-NVFP4 to start chatting
Load model with FastModel
pip install unsloth from unsloth import FastModel model, tokenizer = FastModel.from_pretrained( model_name="unsloth/Qwen3.8-27B-NVFP4", max_seq_length=2048, )
Using this on vllm but excessive tbinking
π 1
2
#13 opened about 4 hours ago
by
crystech
Changelog?
1
#12 opened about 10 hours ago
by
sparse-simian
Does not work with vllm 0.27.1 (latest)
4
#11 opened about 12 hours ago
by
flaviusburca
Tokenizer bug allows only small images
β€οΈ 1
2
#10 opened about 17 hours ago
by
HackyHH
thinking tags
4
#9 opened about 21 hours ago
by
MalcolmMielle
Inferact/Qwen3.8-27B-NVFP4 is on vLLM's official documentation
5
#8 opened 1 day ago
by
pathosethoslogos
Qwen3.8-27B Serving Configs: DGX Spark vLLM NVFP4
π 6
5
#7 opened 1 day ago
by
erdal
Why no NVFP4 GGUF?
ππ 5
2
#6 opened 1 day ago
by
mattzink
What is the target size for NVFP4?
4
#5 opened 1 day ago
by
musa334
Can it be reduced much further?
π 1
1
#4 opened 1 day ago
by
Duonglv
So fast and great. Thank so muchhhh.
1
#3 opened 1 day ago
by
Duonglv
Thank you
4
#1 opened 1 day ago
by
pubdddq23erq