MeetInstruct-0.6B-v1.5 GGUF

GGUF conversions of Ma7ee7/MeetInstruct-0.6B-v1.5.

The model is based on Qwen3-0.6B and uses its ChatML-compatible chat template.

Files

File Size
MeetInstruct-0.6B-v1.5-F16.gguf 1439.4 MiB
MeetInstruct-0.6B-v1.5-Q8_0.gguf 767.5 MiB
MeetInstruct-0.6B-v1.5-Q6_K.gguf 593.9 MiB
MeetInstruct-0.6B-v1.5-Q5_K_M.gguf 525.8 MiB
MeetInstruct-0.6B-v1.5-Q5_K_S.gguf 518.4 MiB
MeetInstruct-0.6B-v1.5-Q4_K_M.gguf 461.8 MiB
MeetInstruct-0.6B-v1.5-Q4_K_S.gguf 449.0 MiB
MeetInstruct-0.6B-v1.5-Q3_K_L.gguf 415.2 MiB
MeetInstruct-0.6B-v1.5-Q3_K_M.gguf 394.8 MiB
MeetInstruct-0.6B-v1.5-Q3_K_S.gguf 371.9 MiB
MeetInstruct-0.6B-v1.5-Q2_K.gguf 331.2 MiB

Recommended general-purpose quantization: MeetInstruct-0.6B-v1.5-Q4_K_M.gguf.

llama.cpp

Run the model with:

llama-cli -m MeetInstruct-0.6B-v1.5-Q4_K_M.gguf -cnv

Context support depends on the underlying checkpoint and available memory. The v1.5 16K checkpoint was trained with examples between approximately 10,000 and 15,900 tokens, but long-context reliability is not guaranteed.

Quantizations

Generated locally using the current llama.cpp quantizer.

Successful quantizations: Q8_0, Q6_K, Q5_K_M, Q5_K_S, Q4_K_M, Q4_K_S, Q3_K_L, Q3_K_M, Q3_K_S, Q2_K.

Downloads last month
249
GGUF
Model size
0.8B params
Architecture
qwen3
Hardware compatibility
Log In to add your hardware

2-bit

3-bit

4-bit

5-bit

6-bit

8-bit

16-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for Ma7ee7/MeetInstruct-0.6B-v1.5-GGUF

Quantized
(2)
this model