LogiShell store Open app

Store › model › LLM

Ling-3.0-tiny-GGUF

by bloomer010 · source Hugging Face · updated 2026-09-20

mit4.5 GB~6 GB RAMsource aliveunlabeled

GGUF conversions of inclusionAI/Ling-3.0-tiny, converted directly from the released BF16 safetensors.

Add to LogiShell Open in the web IDE

The button opens LogiShell with this card; nothing installs from a link by itself. Inside the app the install goes through lsh models install hf:bloomer010/Ling-3.0-tiny-GGUF and its progress lives in the Resource Center.

Source and license

Numbers

Numbers as of 2026-10-11 15:33 UTC, from the source API.

Summary

Summary not ready yet: the numbers are here, the text is not. It is written by the collector through the LogiShell model facade when a provider key is present.

Reviews

No reviews yet. Reviews are written inside LogiShell: open this card in the app.

Files

filequantsize
Ling-3.0-tiny-BF16.ggufBF1615 GB
Ling-3.0-tiny-IQ1_M.ggufIQ1_M1.8 GB
Ling-3.0-tiny-IQ1_S.ggufIQ1_S1.6 GB
Ling-3.0-tiny-IQ2_M.ggufIQ2_M2.5 GB
Ling-3.0-tiny-IQ2_S.ggufIQ2_S2.3 GB
Ling-3.0-tiny-IQ2_XS.ggufIQ2_XS2.3 GB
Ling-3.0-tiny-IQ2_XXS.ggufIQ2_XXS2.1 GB
Ling-3.0-tiny-IQ3_S.ggufIQ3_S3.3 GB
Ling-3.0-tiny-IQ3_XXS.ggufIQ3_XXS2.9 GB
Ling-3.0-tiny-IQ4_XS.ggufIQ4_XS4.0 GB
Ling-3.0-tiny-MXFP4_MOE.gguf4.4 GB
Ling-3.0-tiny-Q1_0.ggufQ1_01.2 GB
Ling-3.0-tiny-Q2_K.ggufQ2_K2.8 GB
Ling-3.0-tiny-Q3_K_M.ggufQ3_K_M3.6 GB
Ling-3.0-tiny-Q3_K_S.ggufQ3_K_S3.3 GB
Ling-3.0-tiny-Q4_0.ggufQ4_04.2 GB
Ling-3.0-tiny-Q4_K_M.ggufQ4_K_M4.5 GB
Ling-3.0-tiny-Q4_K_S.ggufQ4_K_S4.2 GB
Ling-3.0-tiny-Q5_0.ggufQ5_05.1 GB
Ling-3.0-tiny-Q5_K_M.ggufQ5_K_M5.2 GB
Ling-3.0-tiny-Q5_K_S.ggufQ5_K_S5.1 GB
Ling-3.0-tiny-Q6_K.ggufQ6_K6.1 GB
Ling-3.0-tiny-Q8_0.ggufQ8_07.8 GB
Ling-3.0-tiny-UD-Q4_K_XL.ggufQ4_K_XL5.0 GB
Ling-3.0-tiny-UD-Q6_K_XL.ggufQ6_K_XL6.8 GB
Ling-3.0-tiny-UD-Q8_K_XL.ggufQ8_K_XL10 GB
Ling-3.0-tiny-f16.ggufF1615 GB
Ling-3.0-tiny-imatrix.gguf42 MB

From the source README

Ling-3.0-tiny GGUF

GGUF conversions of inclusionAI/Ling-3.0-tiny,
converted directly from the released BF16 safetensors.

## 🦙🚨 llama.cpp 🦙🚨
Consistent agentic use (tool calling, reasoning split) currently requires two unmerged llama.cpp PRs:
- Dedicated Ling parser: #28682 (✅ Merged as of 9/19)
- Invalid UTF-8 handling in the PEG parser: #29161 (✅ Merged as of 9/20)

Without both, tool calls inside an unclosed think block are dropped and some turns fail with a 500.
Will update this note as they merge.

The model does occasionally terminate its response, mid-think, without any sort of closing.
This is inherent in the weights, even at full precision.

To run with `llama-server`:
```bash
llama-server -hf bloomer010/Ling-3.0-tiny-GGUF:Q4_K_M
```

Quant Sizing

For tiny models, precision is especially crucial.

Generally...
Larger files = More precision.
Smaller files = More compression = More slop and misbehavin'.

Use UD-Q8_K_XL for near-full precision performance.

| Quant | Size | your memory |
| --- | ---: | --- |
| BF16 | 15.8 GB | 16 GB+ |
| UD-Q8_K_XL | 11.19 GB | 12 GB+ |
| Q8_0 | 8.41 GB | 10 GB+ |
| UD-Q6_K_XL | 7.27 GB | 8 GB+ |
| Q6_K | 6.50 GB | 8 GB+ |
| Q5_K_M | 5.64 GB | 7 GB+ |
| Q5_K_S | 5.48 GB | 6 GB+ |
| Q5_0 | 5.48 GB | 6 GB+ |
| Q4_K_M | 4.82 GB | 6 GB+ |
| Q4_K_S | 4.55 GB | 6 GB+ |
| Q4_0 | 4.53 GB | 6 GB+ |
| MXFP4_MOE | 4.72 GB | 6 GB+ ¹ |
| IQ4_XS | 4.29 GB | 5 GB+ |
| Q3_K_M | 3.84 GB | 5 GB+ |
| Q3_K_S | 3.51 GB | 5 GB+ |
| IQ3_S | 3.51 GB | 4 GB+ |
| IQ3_XXS | 3.13 GB | 4 GB+ |
| Q2_K | 2.99 GB | 4 GB+ |
| IQ2_M | 2.70 GB | 3 GB+ |
| IQ2_S | 2.48 GB | 3 GB+ |
| IQ2_XS | 2.43 GB | 3 GB+ |
| IQ2_XXS | 2.21 GB | 3 GB+ |
| IQ1_M | 1.93 GB | 3 GB+ |
| IQ1_S | 1.76 GB | 2 GB+ |
| Q1_0 | 1.30 GB | 2 GB+ |

¹ `MXFP4_MOE`…

Card id model:hf:bloomer010/Ling-3.0-tiny-GGUF · collected 2026-10-11 15:33 UTC · JSON