Routed experts in NVIDIA's NVFP4, everything else Q8_0. This mix helps with performace on nvidia gpu.

Downloads last month
19
GGUF
Model size
35B params
Architecture
qwen35moe
Hardware compatibility
Log In to add your hardware

8-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for CompiledThoughts/Qwen3.6-35B-A3B-NVFP4-Q8_0-it

Quantized
(10)
this model