Tensor-0.1-35B-A3B

A Self-postrained NVFP4-quantized 35B MoE model derived from Qwen3.5-35B-A3B.

Overview

Property Value
Base model Qwen3.5-35B-A3B
Architecture Qwen3_5Moe (Mixture-of-Experts)
Quantization NVFP4 (4-bit floating point)
Context length 262,144 tokens
Model size ~17 GB

Usage

from transformers import AutoModelForCausalLM, AutoTokenizer

model = AutoModelForCausalLM.from_pretrained("Akushon/Tensor-0.1-35B-A3B")
tokenizer = AutoTokenizer.from_pretrained("Akushon/Tensor-0.1-35B-A3B")

Note: This model requires NVFP4 quantization support. Use a compatible NVFP4-aware inference backend.

License

This model is released under the same license terms as the base model.

Downloads last month
-
Safetensors
Model size
15B params
Tensor type
BF16
·
F8_E4M3
·
U8
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for Akushon/Tensor-0.1-35B-A3B

Quantized
(285)
this model