Azulta Web 7B - GGUF

GGUF quantizations of youte3/azulta-web-7b, a LoRA fine-tune of Qwen2.5-7B-Instruct merged back into the base model.

Available Quantizations

Quantization File
Q4_K_M azulta-web-7b-Q4_K_M.gguf
Q5_K_M azulta-web-7b-Q5_K_M.gguf

Usage

Use with llama.cpp, ollama, LM Studio, or any GGUF-compatible runtime.

Downloads last month
8
GGUF
Model size
8B params
Architecture
qwen2
Hardware compatibility
Log In to add your hardware

4-bit

5-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for youte3/azulta-web-7b-gguf

Base model

Qwen/Qwen2.5-7B
Quantized
(411)
this model