LogiShell store Open app

Store › model › LLM

Qwen2.5-32B-Instruct-GGUF

by bartowski · source Hugging Face · updated 2024-09-19

apache-2.018 GB~22 GB RAMsource aliveunlabeled

Llamacpp imatrix Quantizations of Qwen2.5-32B-Instruct

Add to LogiShell Open in the web IDE

The button opens LogiShell with this card; nothing installs from a link by itself. Inside the app the install goes through lsh models install hf:bartowski/Qwen2.5-32B-Instruct-GGUF and its progress lives in the Resource Center.

Source and license

Numbers

Numbers as of 2026-10-02 21:00 UTC, from the source API.

Summary

Summary not ready yet: the numbers are here, the text is not. It is written by the collector through the LogiShell model facade when a provider key is present.

Reviews

No reviews yet. Reviews are written inside LogiShell: open this card in the app.

Files

filequantsize
Qwen2.5-32B-Instruct-IQ2_M.ggufIQ2_M10 GB
Qwen2.5-32B-Instruct-IQ2_S.ggufIQ2_S9.7 GB
Qwen2.5-32B-Instruct-IQ2_XS.ggufIQ2_XS9.3 GB
Qwen2.5-32B-Instruct-IQ2_XXS.ggufIQ2_XXS8.4 GB
Qwen2.5-32B-Instruct-IQ3_M.ggufIQ3_M14 GB
Qwen2.5-32B-Instruct-IQ3_XS.ggufIQ3_XS13 GB
Qwen2.5-32B-Instruct-IQ4_XS.ggufIQ4_XS16 GB
Qwen2.5-32B-Instruct-Q2_K.ggufQ2_K11 GB
Qwen2.5-32B-Instruct-Q2_K_L.ggufQ2_K_L12 GB
Qwen2.5-32B-Instruct-Q3_K_L.ggufQ3_K_L16 GB
Qwen2.5-32B-Instruct-Q3_K_M.ggufQ3_K_M15 GB
Qwen2.5-32B-Instruct-Q3_K_S.ggufQ3_K_S13 GB
Qwen2.5-32B-Instruct-Q3_K_XL.ggufQ3_K_XL17 GB
Qwen2.5-32B-Instruct-Q4_0.ggufQ4_017 GB
Qwen2.5-32B-Instruct-Q4_0_4_4.ggufQ4_017 GB
Qwen2.5-32B-Instruct-Q4_0_4_8.ggufQ4_017 GB
Qwen2.5-32B-Instruct-Q4_0_8_8.ggufQ4_017 GB
Qwen2.5-32B-Instruct-Q4_K_L.ggufQ4_K_L19 GB
Qwen2.5-32B-Instruct-Q4_K_M.ggufQ4_K_M18 GB
Qwen2.5-32B-Instruct-Q4_K_S.ggufQ4_K_S17 GB
Qwen2.5-32B-Instruct-Q5_K_L.ggufQ5_K_L22 GB
Qwen2.5-32B-Instruct-Q5_K_M.ggufQ5_K_M22 GB
Qwen2.5-32B-Instruct-Q5_K_S.ggufQ5_K_S21 GB
Qwen2.5-32B-Instruct-Q6_K.ggufQ6_K25 GB
Qwen2.5-32B-Instruct-Q6_K_L.ggufQ6_K_L25 GB
Qwen2.5-32B-Instruct-Q8_0.ggufQ8_032 GB
Qwen2.5-32B-Instruct-f16/Qwen2.5-32B-Instruct-f16-00001-of-00002.ggufF1637 GB
Qwen2.5-32B-Instruct-f16/Qwen2.5-32B-Instruct-f16-00002-of-00002.ggufF1624 GB

From the source README

Llamacpp imatrix Quantizations of Qwen2.5-32B-Instruct

Using llama.cpp release b3772 for quantization.

Original model: https://huggingface.co/Qwen/Qwen2.5-32B-Instruct

All quants made using imatrix option with dataset from here

Run them in LM Studio

Prompt format

system
{system_prompt}
user
{prompt}
assistant

What's new:

Update context length settings and tokenizer

Download a file (not the whole branch) from below:

| Filename | Quant type | File Size | Split | Description |
| -------- | ---------- | --------- | ----- | ----------- |
| Qwen2.5-32B-Instruct-f16.gguf | f16 | 65.54GB | true | Full F16 weights. |
| Qwen2.5-32B-Instruct-Q8_0.gguf | Q8_0 | 34.82GB | false | Extremely high quality, generally unneeded but max available quant. |
| Qwen2.5-32B-Instruct-Q6_K_L.gguf | Q6_K_L | 27.26GB | false | Uses Q8_0 for embed and output weights. Very high quality, near perfect, *recommended*. |
| Qwen2.5-32B-Instruct-Q6_K.gguf | Q6_K | 26.89GB | false | Very high quality, near perfect, *recommended*. |
| Qwen2.5-32B-Instruct-Q5_K_L.gguf | Q5_K_L | 23.74GB | false | Uses Q8_0 for embed and output weights. High quality, *recommended*. |
| Qwen2.5-32B-Instruct-Q5_K_M.gguf | Q5_…

Source: https://huggingface.co/bartowski/Qwen2.5-32B-Instruct-GGUF

Card id model:hf:bartowski/Qwen2.5-32B-Instruct-GGUF · collected 2026-10-02 21:00 UTC · JSON