Store › model › LLM
Ling-3.0-tiny-GGUF
by bloomer010 · source Hugging Face · updated 2026-09-20
mit4.5 GB~6 GB RAMsource aliveunlabeled
GGUF conversions of inclusionAI/Ling-3.0-tiny, converted directly from the released BF16 safetensors.
Add to LogiShell Open in the web IDE
The button opens LogiShell with this card; nothing installs from a link by itself. Inside the app the install goes through lsh models install hf:bloomer010/Ling-3.0-tiny-GGUF and its progress lives in the Resource Center.
Source and license
- Source: https://huggingface.co/bloomer010/Ling-3.0-tiny-GGUF
- License: mit
- Requirements: about 6 GB of RAM, 4.5 GB on disk (estimate: suggested file size × 1.15 + 0.5 GB; a real measurement comes with lsh models). Runs with llama.cpp, ollama.
- Tags:
llama.cppggufbailing_hybridbailingmoe3mixture-of-expertsconversationaltext-generationcustom_codeendpoints_compatible
Numbers
- 265,045 downloads on Hugging Face
- 112 likes
- license mit
- 4.5 GB for Ling-3.0-tiny-Q4_K_M.gguf
- 130,822 stars on ggml-org/llama.cpp
- 2,517 open issues and PRs
- last release v0.6.0 on 2026-10-05
- 261,158 npm downloads a week for node-llama-cpp
- latest node-llama-cpp@3.22.1
Numbers as of 2026-10-11 15:33 UTC, from the source API.
Summary
Summary not ready yet: the numbers are here, the text is not. It is written by the collector through the LogiShell model facade when a provider key is present.
Reviews
No reviews yet. Reviews are written inside LogiShell: open this card in the app.
Files
| file | quant | size |
|---|---|---|
Ling-3.0-tiny-BF16.gguf | BF16 | 15 GB |
Ling-3.0-tiny-IQ1_M.gguf | IQ1_M | 1.8 GB |
Ling-3.0-tiny-IQ1_S.gguf | IQ1_S | 1.6 GB |
Ling-3.0-tiny-IQ2_M.gguf | IQ2_M | 2.5 GB |
Ling-3.0-tiny-IQ2_S.gguf | IQ2_S | 2.3 GB |
Ling-3.0-tiny-IQ2_XS.gguf | IQ2_XS | 2.3 GB |
Ling-3.0-tiny-IQ2_XXS.gguf | IQ2_XXS | 2.1 GB |
Ling-3.0-tiny-IQ3_S.gguf | IQ3_S | 3.3 GB |
Ling-3.0-tiny-IQ3_XXS.gguf | IQ3_XXS | 2.9 GB |
Ling-3.0-tiny-IQ4_XS.gguf | IQ4_XS | 4.0 GB |
Ling-3.0-tiny-MXFP4_MOE.gguf | 4.4 GB | |
Ling-3.0-tiny-Q1_0.gguf | Q1_0 | 1.2 GB |
Ling-3.0-tiny-Q2_K.gguf | Q2_K | 2.8 GB |
Ling-3.0-tiny-Q3_K_M.gguf | Q3_K_M | 3.6 GB |
Ling-3.0-tiny-Q3_K_S.gguf | Q3_K_S | 3.3 GB |
Ling-3.0-tiny-Q4_0.gguf | Q4_0 | 4.2 GB |
Ling-3.0-tiny-Q4_K_M.gguf | Q4_K_M | 4.5 GB |
Ling-3.0-tiny-Q4_K_S.gguf | Q4_K_S | 4.2 GB |
Ling-3.0-tiny-Q5_0.gguf | Q5_0 | 5.1 GB |
Ling-3.0-tiny-Q5_K_M.gguf | Q5_K_M | 5.2 GB |
Ling-3.0-tiny-Q5_K_S.gguf | Q5_K_S | 5.1 GB |
Ling-3.0-tiny-Q6_K.gguf | Q6_K | 6.1 GB |
Ling-3.0-tiny-Q8_0.gguf | Q8_0 | 7.8 GB |
Ling-3.0-tiny-UD-Q4_K_XL.gguf | Q4_K_XL | 5.0 GB |
Ling-3.0-tiny-UD-Q6_K_XL.gguf | Q6_K_XL | 6.8 GB |
Ling-3.0-tiny-UD-Q8_K_XL.gguf | Q8_K_XL | 10 GB |
Ling-3.0-tiny-f16.gguf | F16 | 15 GB |
Ling-3.0-tiny-imatrix.gguf | 42 MB |
From the source README
Ling-3.0-tiny GGUF
GGUF conversions of inclusionAI/Ling-3.0-tiny,
converted directly from the released BF16 safetensors.
## 🦙🚨 llama.cpp 🦙🚨
Consistent agentic use (tool calling, reasoning split) currently requires two unmerged llama.cpp PRs:
- Dedicated Ling parser: #28682 (✅ Merged as of 9/19)
- Invalid UTF-8 handling in the PEG parser: #29161 (✅ Merged as of 9/20)
Without both, tool calls inside an unclosed think block are dropped and some turns fail with a 500.
Will update this note as they merge.
The model does occasionally terminate its response, mid-think, without any sort of closing.
This is inherent in the weights, even at full precision.
To run with `llama-server`:
```bash
llama-server -hf bloomer010/Ling-3.0-tiny-GGUF:Q4_K_M
```
Quant Sizing
For tiny models, precision is especially crucial.
Generally...
Larger files = More precision.
Smaller files = More compression = More slop and misbehavin'.
Use UD-Q8_K_XL for near-full precision performance.
| Quant | Size | your memory |
| --- | ---: | --- |
| BF16 | 15.8 GB | 16 GB+ |
| UD-Q8_K_XL | 11.19 GB | 12 GB+ |
| Q8_0 | 8.41 GB | 10 GB+ |
| UD-Q6_K_XL | 7.27 GB | 8 GB+ |
| Q6_K | 6.50 GB | 8 GB+ |
| Q5_K_M | 5.64 GB | 7 GB+ |
| Q5_K_S | 5.48 GB | 6 GB+ |
| Q5_0 | 5.48 GB | 6 GB+ |
| Q4_K_M | 4.82 GB | 6 GB+ |
| Q4_K_S | 4.55 GB | 6 GB+ |
| Q4_0 | 4.53 GB | 6 GB+ |
| MXFP4_MOE | 4.72 GB | 6 GB+ ¹ |
| IQ4_XS | 4.29 GB | 5 GB+ |
| Q3_K_M | 3.84 GB | 5 GB+ |
| Q3_K_S | 3.51 GB | 5 GB+ |
| IQ3_S | 3.51 GB | 4 GB+ |
| IQ3_XXS | 3.13 GB | 4 GB+ |
| Q2_K | 2.99 GB | 4 GB+ |
| IQ2_M | 2.70 GB | 3 GB+ |
| IQ2_S | 2.48 GB | 3 GB+ |
| IQ2_XS | 2.43 GB | 3 GB+ |
| IQ2_XXS | 2.21 GB | 3 GB+ |
| IQ1_M | 1.93 GB | 3 GB+ |
| IQ1_S | 1.76 GB | 2 GB+ |
| Q1_0 | 1.30 GB | 2 GB+ |
¹ `MXFP4_MOE`…
Card id model:hf:bloomer010/Ling-3.0-tiny-GGUF · collected 2026-10-11 15:33 UTC · JSON