llm-qwen

Qwen 3.5 summary models for Tetro, the local meeting transcriber by Vana Labs, in GGUF format. They run fully on your computer.

This repository is a pinned mirror, so Tetro's downloads don't depend on third-party hosts. The files are unchanged copies.

File Tetro model ID Source
Qwen3.5-2B-Q4_K_M.gguf qwen3.5:2b unsloth/Qwen3.5-2B-GGUF
Qwen3.5-4B-Q4_K_M.gguf qwen3.5:4b unsloth/Qwen3.5-4B-GGUF

Models by the Qwen team, Alibaba Cloud. Apache-2.0.

Source commits: unsloth/Qwen3.5-2B-GGUF f6d5376be1edb4d416d56da11e5397a961aca8ae, unsloth/Qwen3.5-4B-GGUF e87f176479d0855a907a41277aca2f8ee7a09523.

Downloads last month
10
GGUF
Model size
2B params
Architecture
qwen35
Hardware compatibility
Log In to add your hardware

4-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for Vana-Labs/llm-qwen

Finetuned
Qwen/Qwen3.5-2B
Quantized
(212)
this model