Qwen3.5-122B-A10B

BaseRT .base builds of Qwen/Qwen3.5-122B-A10B for fast local inference on Apple Silicon (Metal) and NVIDIA GB10 (CUDA).

Files

File Precision Size
Qwen3.5-122B-A10B-Q4.base Q4 69 GB
Qwen3.5-122B-A10B-Q8.base Q8 126 GB

Usage

curl -LsSf https://basecompute.co/install.sh | sh
basert pull basecompute/Qwen3.5-122B-A10B
basert chat basecompute/Qwen3.5-122B-A10B

Built and measured at baseRT commit hf.

Released under the apache-2.0 license, inherited from the base model.

Downloads last month

-

Downloads are not tracked for this model. How to track
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for basecompute/Qwen3.5-122B-A10B

Finetuned
(51)
this model