Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up

unsloth
/
Qwen3.8-27B-NVFP4

Safetensors
qwen3_5
unsloth
compressed-tensors
Model card Files Files and versions
xet
Community
13

Instructions to use unsloth/Qwen3.8-27B-NVFP4 with libraries, inference providers, notebooks, and local apps. Follow these links to get started.

  • Local Apps Settings
  • Unsloth Studio

    How to use unsloth/Qwen3.8-27B-NVFP4 with Unsloth Studio:

    Install Unsloth Studio (macOS, Linux, WSL)
    curl -fsSL https://unsloth.ai/install.sh | sh
    # Run unsloth studio
    unsloth studio -H 0.0.0.0 -p 8888
    # Then open http://localhost:8888 in your browser
    # Search for unsloth/Qwen3.8-27B-NVFP4 to start chatting
    Install Unsloth Studio (Windows)
    irm https://unsloth.ai/install.ps1 | iex
    # Run unsloth studio
    unsloth studio -H 0.0.0.0 -p 8888
    # Then open http://localhost:8888 in your browser
    # Search for unsloth/Qwen3.8-27B-NVFP4 to start chatting
    Using HuggingFace Spaces for Unsloth
    # No setup required
    # Open https://huggingface.co/spaces/unsloth/studio in your browser
    # Search for unsloth/Qwen3.8-27B-NVFP4 to start chatting
    Load model with FastModel
    pip install unsloth
    from unsloth import FastModel
    model, tokenizer = FastModel.from_pretrained(
        model_name="unsloth/Qwen3.8-27B-NVFP4",
        max_seq_length=2048,
    )
New discussion
Resources
  • PR & discussions documentation
  • Code of Conduct
  • Hub documentation

Using this on vllm but excessive tbinking

πŸ‘€ 1
2
#13 opened about 4 hours ago by
crystech

Changelog?

1
#12 opened about 10 hours ago by
sparse-simian

Does not work with vllm 0.27.1 (latest)

4
#11 opened about 12 hours ago by
flaviusburca

Tokenizer bug allows only small images

❀️ 1
2
#10 opened about 17 hours ago by
HackyHH

thinking tags

4
#9 opened about 21 hours ago by
MalcolmMielle

Inferact/Qwen3.8-27B-NVFP4 is on vLLM's official documentation

5
#8 opened 1 day ago by
pathosethoslogos

Qwen3.8-27B Serving Configs: DGX Spark vLLM NVFP4

πŸš€ 6
5
#7 opened 1 day ago by
erdal

Why no NVFP4 GGUF?

πŸ‘πŸ‘€ 5
2
#6 opened 1 day ago by
mattzink

What is the target size for NVFP4?

4
#5 opened 1 day ago by
musa334

Can it be reduced much further?

πŸ‘ 1
1
#4 opened 1 day ago by
Duonglv

So fast and great. Thank so muchhhh.

1
#3 opened 1 day ago by
Duonglv

How ?

3
#2 opened 1 day ago by
jbourny

Thank you

4
#1 opened 1 day ago by
pubdddq23erq
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs