PREVIEW 1.5 — GGUF quants for testing. Final release will follow after think-SFT, native MTP and updated quants.

CobrIX-1.5-preview-Coder-Flash-33B-A13B-GGUF

GGUF quantizations of CobrIX/CobrIX-1.5-preview-Coder-Flash-33B-A13B — MoE decoder (~33B total / ~13B active per token), DPO-aligned + SFT-reinforced for code and cybersecurity (PT/EN).

Quants

File Size VRAM/RAM needed (approx.)
CobrIX-1.5-preview-Coder-Flash-33B-A13B-Q5_K_M.gguf 22,0 GB ~24 GB
CobrIX-1.5-preview-Coder-Flash-33B-A13B-Q4_K_M.gguf 18,9 GB ~20 GB

Q5_K_M = best quality here; Q4_K_M = smaller/faster with minimal loss. ChatML template embedded (<|im_start|>/<|im_end|>), stop at <|im_end|>.

Usage (Ollama)

FROM ./CobrIX-1.5-preview-Coder-Flash-33B-A13B-Q5_K_M.gguf

TEMPLATE """{{- if .System }}<|im_start|>system
{{ .System }}<|im_end|>
{{ end }}<|im_start|>user
{{ .Prompt }}<|im_end|>
<|im_start|>assistant
"""

PARAMETER stop "<|im_end|>"
PARAMETER num_ctx 8192
ollama create cobrix15-flash-q5 -f Modelfile
ollama run cobrix15-flash-q5 "Explique SQL injection e como mitigar."

Usage (llama.cpp)

llama-cli -m CobrIX-1.5-preview-Coder-Flash-33B-A13B-Q5_K_M.gguf \
  --jinja -p "<|im_start|>user\nExplique SQL injection<|im_end|>\n<|im_start|>assistant\n" -n 300

🌐 CobrIX Coder and CobrIX Code Models — Early Access

CobrIX Coder is the AI ecosystem that runs and gives you instant access to our models — including CobrIX-1.0-Coder-Flash-MoE, CobrIX-1.0-Coder-Full-MoE, and the brand-new 1.5 preview family.

✨ Try CobrIX Coder Live Now

https://cobrix.vercel.app/coder

Explore the platform, test the models in real time, and follow the evolution of CobrIX Code as it happens.

⚡ Flash First — Priority Access Coming Soon!

We're gearing up for the big public launch! As soon as our waitlist hits a strong number, we'll unlock the Flash heroes firstCobrIX-1.0-Coder-Flash and CobrIX-1.5-Coder-Flash — as soon as everything is polished and ready.

Early sign-ups get priority access. Secure your spot today and be among the very first to experience the next generation! 🌟

📋 Interest List — Join the Waitlist

Want to be the first to know when the CobrIX AI ecosystem and CobrIX Code go fully public?

👉 Visit: https://cobrix.vercel.app/coder

Registered users will receive an email the moment CobrIX Code becomes available.

We currently have the funds to host CobrIX-1.0-Coder-Flash-33B-A13B on an RTX 6000 Ada Generation (48GB VRAM), supporting 50–100 concurrent users.

🚀 Join the waitlist now and be first in line when CobrIX Code launches!

Feedback — help make it better

Found a flaw or have an idea? 📧 suporte.cobrix@gmail.com — send the prompt, the output, what you expected, and how to improve it.

Donations for Infrastructure

Support open-source CobrIX development:

Bitcoin (BTC)

bc1q8mu8fjak4y84qj4dlk8pu4d3zhknm92zra4r4m

Ethereum (ETH / ERC-20)

0x8D9187dEa0a77390ef668361cd5b236DE54af2BB

Solana (SOL)

GQR2jZnWuWP1c3dbuz4mC7ZnyveacBKy63q8qf9nj8bp

CobrIX — Open AI infrastructure and custom model research.

Downloads last month
145
GGUF
Model size
33B params
Architecture
qwen35moe
Hardware compatibility
Log In to add your hardware

4-bit

5-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for CobrIX/CobrIX-1.5-preview-Coder-Flash-33B-A13B-GGUF

Collection including CobrIX/CobrIX-1.5-preview-Coder-Flash-33B-A13B-GGUF