Experimental version using https://github.com/ggml-org/llama.cpp/pull/26185 For testing purposes "Full quality" as the original weights are mostly in mxfp4

Downloads last month
-
GGUF
Model size
2.8T params
Architecture
kimi-k3
Hardware compatibility
Log In to add your hardware

We're not able to determine the quantization variants.

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for bullerwins/Kimi-K3-GGUF

Quantized
(18)
this model