{"v":1,"id":"model:hf:Qwen/Qwen2.5-1.5B-Instruct-GGUF","slug":"model-qwen-qwen2-5-1-5b-instruct-gguf","kind":"model","category":"llm","title":"Qwen2.5-1.5B-Instruct-GGUF","summary":"Qwen2.5 is the latest series of Qwen large language models. For Qwen2.5, we release a number of base language models and instruction-tuned language models ranging from 0.5 to 72 billion parameters. Q…","source":{"provider":"hf","ref":"Qwen/Qwen2.5-1.5B-Instruct-GGUF","url":"https://huggingface.co/Qwen/Qwen2.5-1.5B-Instruct-GGUF","rev":"91cad51170dc346986eccefdc2dd33a9da36ead9","fetchedAt":"2026-10-09T19:01:17.158Z","etag":"W/\"22cf-p8naNj4xJKKogc27ME9iaRJt1/4\""},"author":{"name":"Qwen","url":"https://huggingface.co/Qwen"},"license":{"spdx":"apache-2.0","raw":"apache-2.0","url":"https://huggingface.co/Qwen/Qwen2.5-1.5B-Instruct-GGUF/blob/main/LICENSE","open":true},"metrics":{"downloads":301376,"downloadsWeek":297971,"likes":170,"stars":27659,"openIssues":68,"pushedAt":"2026-01-09T03:05:47Z","takenAt":"2026-10-09T19:01:17.158Z"},"tags":["gguf","chat","text-generation","en","endpoints_compatible","conversational"],"pipeline":"text-generation","links":{"github":"QwenLM/Qwen3","npm":"node-llama-cpp"},"updatedAt":"2024-09-20T06:31:38.000Z","collectedAt":"2026-10-09T19:01:17.158Z","review":{"numbers":["301,376 downloads on Hugging Face","170 likes","license apache-2.0","1.0 GB for qwen2.5-1.5b-instruct-q4_k_m.gguf","27,659 stars on QwenLM/Qwen3","68 open issues and PRs","297,971 npm downloads a week for node-llama-cpp","latest node-llama-cpp@3.22.1"],"log":null},"trust":"unlabeled","health":{"status":"alive","checkedAt":"2026-10-09T19:01:17.158Z","http":200},"description":"# Qwen2.5-1.5B-Instruct-GGUF\n\n## Introduction\n\nQwen2.5 is the latest series of Qwen large language models. For Qwen2.5, we release a number of base language models and instruction-tuned language models ranging from 0.5 to 72 billion parameters. Qwen2.5 brings the following improvements upon Qwen2:\n\n- Significantly **more knowledge** and has greatly improved capabilities in **coding** and **mathematics**, thanks to our specialized expert models in these domains.\n- Significant improvements in **instruction following**, **generating long texts** (over 8K tokens), **understanding structured data** (e.g, tables), and **generating structured outputs** especially JSON. **More resilient to the diversity of system prompts**, enhancing role-play implementation and condition-setting for chatbots.\n- **Long-context Support** up to 128K tokens and can generate up to 8K tokens.\n- **Multilingual support** for over 29 languages, including Chinese, English, French, Spanish, Portuguese, German, Italian, Russian, Japanese, Korean, Vietnamese, Thai, Arabic, and more.\n\n**This repo contains the instruction-tuned 1.5B Qwen2.5 model in the GGUF Format**, which has the following features:\n- Type: Causal Language Models\n- Training Stage: Pretraining & Post-training\n- Architecture: transformers with RoPE, SwiGLU, RMSNorm, Attention QKV bias and tied word embeddings\n- Number of Parameters: 1.54B\n- Number of Paramaters (Non-Embedding): 1.31B\n- Number of Layers: 28\n- Number of Attention Heads (GQA): 12 for Q and 2 for KV\n- Context Length: Full 32,768 tokens and generation 8192 tokens\n- Quantization: q2_K, q3_K_M, q4_0, q4_K_M, q5_0, q5_K_M, q6_K, q8_0\n\nFor more details, please refer to our [blog](https://qwenlm.github.io/blog/qwen2.5/), [GitHub](https://github.com/QwenLM/Qwen2.5), and [Documentation](https://qwen.readthedocs.io/en/latest/).\n\n## Quickstart\n\nCheck out our [llama.cpp documentation](https://qwen.readthedocs.io/en/latest…\n\nSource: https://huggingface.co/Qwen/Qwen2.5-1.5B-Instruct-GGUF","install":{"kind":"model","hfId":"Qwen/Qwen2.5-1.5B-Instruct-GGUF","gated":false,"format":"gguf","files":[{"name":"qwen2.5-1.5b-instruct-fp16.gguf","size":3560416288,"sha256":"fc89e330deb3fd8fa560f1c0f35a1e2b8da96d59e13445559ed190307a6f5649"},{"name":"qwen2.5-1.5b-instruct-q2_k.gguf","size":752880160,"quant":"Q2_K","sha256":"5ede348e91ce1e7a330926ec5b202c27b864d065149dc463257fde1f98865b3a"},{"name":"qwen2.5-1.5b-instruct-q3_k_m.gguf","size":924455968,"quant":"Q3_K_M","sha256":"58cb5c05ecef48e82961f1a2be6544145ea26136f69dddda4bbbd092f0e4b993"},{"name":"qwen2.5-1.5b-instruct-q4_0.gguf","size":1066227232,"quant":"Q4_0","sha256":"dcd819ff094852c38faba6873d8ff0c9d51eadb2844539e52042ae5d647bbfdb"},{"name":"qwen2.5-1.5b-instruct-q4_k_m.gguf","size":1117320736,"quant":"Q4_K_M","sha256":"6a1a2eb6d15622bf3c96857206351ba97e1af16c30d7a74ee38970e434e9407e"},{"name":"qwen2.5-1.5b-instruct-q5_0.gguf","size":1259173408,"quant":"Q5_0","sha256":"a579334a7b19838b19f7855252b6bc08b012b46e338cf1494a88e77509cfe4d9"},{"name":"qwen2.5-1.5b-instruct-q5_k_m.gguf","size":1285494304,"quant":"Q5_K_M","sha256":"b46661073c18e5b56a41fa320975f866a00def1ff08feef4718e013258896f8c"},{"name":"qwen2.5-1.5b-instruct-q6_k.gguf","size":1464178720,"quant":"Q6_K","sha256":"e16d94f3b1eb243f6f6be9eee51090ef5dfd741324394fd5b6e0e425c33df5c7"},{"name":"qwen2.5-1.5b-instruct-q8_0.gguf","size":1894532128,"quant":"Q8_0","sha256":"d7efb072e7724d25048a4fda0a3e10b04bdef5d06b1403a1c93bd9f1240a63c8"}],"totalBytes":13324678944,"suggestedFile":"qwen2.5-1.5b-instruct-q4_k_m.gguf","requirements":{"ramGb":2,"diskBytes":1117320736,"note":"estimate: suggested file size × 1.15 + 0.5 GB; a real measurement comes with lsh models"},"runWith":["llama.cpp","ollama"],"command":"lsh models install hf:Qwen/Qwen2.5-1.5B-Instruct-GGUF"}}