'Make knowledge free for everyone'

Experimental!

at this point i'm not suggeting this to be used for anyting else than test!

Based on https://github.com/ggml-org/llama.cpp/pull/26185

Q2_K | text-only | 962200.03 MiB (2.90 BPW)

Simple zeroshot demo

k3_early_demo-ezgif.com-video-to-gif-converter

Quantized version of: moonshotai/Kimi-K3 Buy Me a Coffee at ko-fi.com

Downloads last month
-
GGUF
Model size
2.8T params
Architecture
kimi-k3
Hardware compatibility
Log In to add your hardware

2-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for DevQuasar/moonshotai.Kimi-K3-GGUF

Quantized
(16)
this model

Collection including DevQuasar/moonshotai.Kimi-K3-GGUF