Kimi-K2.7-Code GGUF โ€” Quantized by BatiAI

BatiFlow moonshot MoE

The coding upgrade to Kimi K2.6 โ€” +21.8% on Kimi Code Bench v2, running on a 512GB Mac Studio. IQ3_XXS / IQ4_XS GGUF of moonshotai/Kimi-K2.7-Code (1T total / 32.6B active MoE, DeepSeek-V3-family architecture). Quantized directly from official Moonshot weights โ€” code+multilingual imatrix, BatiAI-signed.

๐Ÿ“ฆ Quantizations

Quant Size Shards Target
IQ3_XXS 394 GB (GiB: 367) 10 M3 Ultra 512GB Mac Studio
IQ4_XS 546 GB (GiB: 509) 13 512GB+ / multi-node / server

Both built from official weights via a Q8_0 intermediate, quantized with a code + EN + KO + ZH imatrix (included: Kimi-K2.7-Code-imatrix.dat). Text-only (the vision tower of the K2.5-family checkpoint is not included; same as other K2 GGUFs).

โœ… Verified (this build, IQ3_XXS) โ€” captured greedy runs:

  • Math: 127+58 โ†’ 185 (clean reasoning trace)
  • Korean: ์„œ์šธ ์†Œ๊ฐœ + ๊น€์น˜ยท๋น„๋น”๋ฐฅยท๋ถˆ๊ณ ๊ธฐ ๊ฐ ํ•œ ๋ฌธ์žฅ โ€” fluent, zero token-mixing or loops
  • Tool-call: {"tool":"get_weather","args":{"city":"๋ถ€์‚ฐ"}} โ€” exact JSON

๐Ÿš€ Usage (llama.cpp โ€” mainline, no fork needed)

hf download batiai/Kimi-K2.7-Code-GGUF "Kimi-K2.7-Code-IQ3_XXS-*.gguf" --local-dir ./k27

# llama.cpp auto-loads all shards from the first one
./llama-cli -m ./k27/Kimi-K2.7-Code-IQ3_XXS-00001-of-00010.gguf -ngl 99 -c 16384 \
  -p "Refactor this function and add tests."

Recommended sampling (Moonshot): --temp 1.0 --top-p 0.95 (thinking mode). Architecture is deepseek2 โ€” supported by mainline llama.cpp out of the box. Ollama tags (batiai/kimi-k2.7-code) follow shortly.

๐Ÿ“œ License

Modified MIT (Moonshot) โ€” commercial use permitted; products exceeding 100M MAU / $20M monthly revenue must display "Kimi K2.7" attribution. Full text at the base model repo. Quantized weights redistributed under the same terms.

โœจ What BatiAI did

  • Direct from official Moonshot weights (never a re-quant of third-party GGUFs)
  • Q8_0 intermediate + diverse imatrix (code/EN/KO/ZH) for balanced fidelity
  • Verified: load โœ… ยท math โœ… ยท Korean โœ… ยท tool-call JSON โœ… โ€” BatiAI metadata-signed

โ€” BatiAI ยท on-device frontier AI ยท https://flow.bati.ai

Downloads last month
55
GGUF
Model size
1T params
Architecture
deepseek2
Hardware compatibility
Log In to add your hardware

3-bit

4-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for batiai/Kimi-K2.7-Code-GGUF

Quantized
(29)
this model

Collection including batiai/Kimi-K2.7-Code-GGUF