Store › model › LLM
Kwaipilot_KAT-Coder-V2.5-Dev-GGUF
by bartowski · source Hugging Face · updated 2026-07-24
apache-2.020 GB~24 GB RAMsource aliveunlabeled
Llamacpp imatrix Quantizations of KAT-Coder-V2.5-Dev by Kwaipilot
Add to LogiShell Open in the web IDE
The button opens LogiShell with this card; nothing installs from a link by itself. Inside the app the install goes through lsh models install hf:bartowski/Kwaipilot_KAT-Coder-V2.5-Dev-GGUF and its progress lives in the Resource Center.
Source and license
- Source: https://huggingface.co/bartowski/Kwaipilot_KAT-Coder-V2.5-Dev-GGUF
- License: apache-2.0
- Requirements: about 24 GB of RAM, 20 GB on disk (estimate: suggested file size × 1.15 + 0.5 GB; a real measurement comes with lsh models). Runs with llama.cpp, ollama.
- Tags:
ggufcodeagentagentic-codingmoecodingtext-generationenzhendpoints_compatibleimatrixconversational
Numbers
- 313,246 downloads on Hugging Face
- 138 likes
- license apache-2.0
- 20 GB for Kwaipilot_KAT-Coder-V2.5-Dev-Q4_K_M.gguf
- 327,838 npm downloads a week for node-llama-cpp
- latest node-llama-cpp@3.22.1
Numbers as of 2026-10-02 21:01 UTC, from the source API.
Summary
Summary not ready yet: the numbers are here, the text is not. It is written by the collector through the LogiShell model facade when a provider key is present.
Reviews
No reviews yet. Reviews are written inside LogiShell: open this card in the app.
Files
| file | quant | size |
|---|---|---|
Kwaipilot_KAT-Coder-V2.5-Dev-IQ2_M.gguf | IQ2_M | 11 GB |
Kwaipilot_KAT-Coder-V2.5-Dev-IQ2_S.gguf | IQ2_S | 10 GB |
Kwaipilot_KAT-Coder-V2.5-Dev-IQ2_XS.gguf | IQ2_XS | 10 GB |
Kwaipilot_KAT-Coder-V2.5-Dev-IQ2_XXS.gguf | IQ2_XXS | 9.1 GB |
Kwaipilot_KAT-Coder-V2.5-Dev-IQ3_M.gguf | IQ3_M | 16 GB |
Kwaipilot_KAT-Coder-V2.5-Dev-IQ3_XS.gguf | IQ3_XS | 15 GB |
Kwaipilot_KAT-Coder-V2.5-Dev-IQ3_XXS.gguf | IQ3_XXS | 14 GB |
Kwaipilot_KAT-Coder-V2.5-Dev-IQ4_NL.gguf | IQ4_NL | 18 GB |
Kwaipilot_KAT-Coder-V2.5-Dev-IQ4_XS.gguf | IQ4_XS | 18 GB |
Kwaipilot_KAT-Coder-V2.5-Dev-Q2_K.gguf | Q2_K | 12 GB |
Kwaipilot_KAT-Coder-V2.5-Dev-Q2_K_L.gguf | Q2_K_L | 12 GB |
Kwaipilot_KAT-Coder-V2.5-Dev-Q3_K_L.gguf | Q3_K_L | 16 GB |
Kwaipilot_KAT-Coder-V2.5-Dev-Q3_K_M.gguf | Q3_K_M | 15 GB |
Kwaipilot_KAT-Coder-V2.5-Dev-Q3_K_S.gguf | Q3_K_S | 14 GB |
Kwaipilot_KAT-Coder-V2.5-Dev-Q3_K_XL.gguf | Q3_K_XL | 16 GB |
Kwaipilot_KAT-Coder-V2.5-Dev-Q4_0.gguf | Q4_0 | 19 GB |
Kwaipilot_KAT-Coder-V2.5-Dev-Q4_1.gguf | Q4_1 | 20 GB |
Kwaipilot_KAT-Coder-V2.5-Dev-Q4_K_L.gguf | Q4_K_L | 20 GB |
Kwaipilot_KAT-Coder-V2.5-Dev-Q4_K_M.gguf | Q4_K_M | 20 GB |
Kwaipilot_KAT-Coder-V2.5-Dev-Q4_K_S.gguf | Q4_K_S | 19 GB |
Kwaipilot_KAT-Coder-V2.5-Dev-Q5_K_L.gguf | Q5_K_L | 24 GB |
Kwaipilot_KAT-Coder-V2.5-Dev-Q5_K_M.gguf | Q5_K_M | 23 GB |
Kwaipilot_KAT-Coder-V2.5-Dev-Q5_K_S.gguf | Q5_K_S | 22 GB |
Kwaipilot_KAT-Coder-V2.5-Dev-Q6_K.gguf | Q6_K | 28 GB |
Kwaipilot_KAT-Coder-V2.5-Dev-Q6_K_L.gguf | Q6_K_L | 28 GB |
Kwaipilot_KAT-Coder-V2.5-Dev-Q8_0.gguf | Q8_0 | 34 GB |
Kwaipilot_KAT-Coder-V2.5-Dev-bf16/Kwaipilot_KAT-Coder-V2.5-Dev-bf16-00001-of-00002.gguf | BF16 | 37 GB |
Kwaipilot_KAT-Coder-V2.5-Dev-bf16/Kwaipilot_KAT-Coder-V2.5-Dev-bf16-00002-of-00002.gguf | BF16 | 28 GB |
Kwaipilot_KAT-Coder-V2.5-Dev-imatrix.gguf | 183 MB |
From the source README
Llamacpp imatrix Quantizations of KAT-Coder-V2.5-Dev by Kwaipilot
Using llama.cpp release b10087 for quantization.
Original model: https://huggingface.co/Kwaipilot/KAT-Coder-V2.5-Dev
All quants made using imatrix option with dataset from here
Run them in your choice of tools:
- llama.cpp
- ramalama
- LM Studio
- koboldcpp
- Jan AI
- Text Generation Web UI
- LoLLMs
- Atomic Chat
Note: if it's a newly supported model, you may need to wait for an update from the developers.
Prompt format
system
{system_prompt}
user
{prompt}
assistant
Download a file (not the whole branch) from below:
| Filename | Quant type | File Size | Split | Description |
| -------- | ---------- | --------- | ----- | ----------- |
| Kwaipilot_KAT-Coder-V2.5-Dev-bf16.gguf | bf16 | 69.38GB | true | Full BF16 weights. |
| Kwaipilot_KAT-Coder-V2.5-Dev-Q8_0.gguf | Q8_0 | 36.91GB | false | Extremely high quality, generally unneeded but max available quant. |
| Kwaipilot_KAT-Coder-V2.5-Dev-Q6_K_L.gguf | Q6_K_L | 30.30GB | false | Uses Q8_0 for embed and output weights. Very high quality, near perfect, *recommended*. |
| [Kwaipilot_KAT-Coder-V2.5-Dev-Q6_K.gguf](https://huggingface.co/bartowski/Kwaipil…
Source: https://huggingface.co/bartowski/Kwaipilot_KAT-Coder-V2.5-Dev-GGUF
Card id model:hf:bartowski/Kwaipilot_KAT-Coder-V2.5-Dev-GGUF · collected 2026-10-02 21:01 UTC · JSON