{"v":1,"id":"model:hf:MaziyarPanahi/Qwen3-0.6B-GGUF","slug":"model-maziyarpanahi-qwen3-0-6b-gguf","kind":"model","category":"llm","title":"Qwen3-0.6B-GGUF","summary":"MaziyarPanahi/Qwen3-0.6B-GGUF - Model creator: Qwen - Original model: Qwen/Qwen3-0.6B","source":{"provider":"hf","ref":"MaziyarPanahi/Qwen3-0.6B-GGUF","url":"https://huggingface.co/MaziyarPanahi/Qwen3-0.6B-GGUF","rev":"16d75108d73a476af91a4f6df4cd77e854b42d04","fetchedAt":"2026-10-02T21:01:11.604Z","etag":"W/\"1d53-N5mSz0NmIgktgWiQfN1Bk5YpKPg\""},"author":{"name":"MaziyarPanahi","url":"https://huggingface.co/MaziyarPanahi"},"license":{"spdx":null,"raw":null,"open":null,"note":"the source did not name a license"},"metrics":{"downloads":289113,"downloadsWeek":327838,"likes":15,"takenAt":"2026-10-02T21:01:11.604Z"},"tags":["gguf","mistral","quantized","2-bit","3-bit","4-bit","5-bit","6-bit","8-bit","GGUF","text-generation","conversational"],"pipeline":"text-generation","links":{"github":"QwenLM/Qwen3","npm":"node-llama-cpp"},"updatedAt":"2025-04-28T21:05:51.000Z","collectedAt":"2026-10-02T21:01:11.604Z","review":{"numbers":["289,113 downloads on Hugging Face","15 likes","license not named","0.5 GB for Qwen3-0.6B.Q4_K_M.gguf","327,838 npm downloads a week for node-llama-cpp","latest node-llama-cpp@3.22.1"],"log":null},"trust":"unlabeled","health":{"status":"alive","checkedAt":"2026-10-02T21:01:11.604Z","http":200},"description":"# [MaziyarPanahi/Qwen3-0.6B-GGUF](https://huggingface.co/MaziyarPanahi/Qwen3-0.6B-GGUF)\n- Model creator: [Qwen](https://huggingface.co/Qwen)\n- Original model: [Qwen/Qwen3-0.6B](https://huggingface.co/Qwen/Qwen3-0.6B)\n\n## Description\n[MaziyarPanahi/Qwen3-0.6B-GGUF](https://huggingface.co/MaziyarPanahi/Qwen3-0.6B-GGUF) contains GGUF format model files for [Qwen/Qwen3-0.6B](https://huggingface.co/Qwen/Qwen3-0.6B).\n\n### About GGUF\n\nGGUF is a new format introduced by the llama.cpp team on August 21st 2023. It is a replacement for GGML, which is no longer supported by llama.cpp.\n\nHere is an incomplete list of clients and libraries that are known to support GGUF:\n\n* [llama.cpp](https://github.com/ggerganov/llama.cpp). The source project for GGUF. Offers a CLI and a server option.\n* [llama-cpp-python](https://github.com/abetlen/llama-cpp-python), a Python library with GPU accel, LangChain support, and OpenAI-compatible API server.\n* [LM Studio](https://lmstudio.ai/), an easy-to-use and powerful local GUI for Windows and macOS (Silicon), with GPU acceleration. Linux available, in beta as of 27/11/2023.\n* [text-generation-webui](https://github.com/oobabooga/text-generation-webui), the most widely used web UI, with many features and powerful extensions. Supports GPU acceleration.\n* [KoboldCpp](https://github.com/LostRuins/koboldcpp), a fully featured web UI, with GPU accel across all platforms and GPU architectures. Especially good for story telling.\n* [GPT4All](https://gpt4all.io/index.html), a free and open source local running GUI, supporting Windows, Linux and macOS with full GPU accel.\n* [LoLLMS Web UI](https://github.com/ParisNeo/lollms-webui), a great web UI with many interesting and unique features, including a full model library for easy model selection.\n* [Faraday.dev](https://faraday.dev/), an attractive and easy to use character-based chat GUI for Windows and macOS (both Silicon and Intel), with GPU acc…\n\nSource: https://huggingface.co/MaziyarPanahi/Qwen3-0.6B-GGUF","install":{"kind":"model","hfId":"MaziyarPanahi/Qwen3-0.6B-GGUF","gated":false,"format":"gguf","files":[{"name":"Qwen3-0.6B.Q2_K.gguf","size":347288704,"quant":"Q2_K","sha256":"4af452189170d7319f135e1eda6f15319681e1ec296200bd59cabcd6b5a27121"},{"name":"Qwen3-0.6B.Q3_K_L.gguf","size":435343488,"quant":"Q3_K_L","sha256":"56af791410421d874f5321f0a2f4d21209ae03596715ddf62262aadbbe84fff6"},{"name":"Qwen3-0.6B.Q3_K_M.gguf","size":413978752,"quant":"Q3_K_M","sha256":"aea2609857f7eddabb22486119a2d8a6416f020bfd7f8fb06c3a7b2db1325811"},{"name":"Qwen3-0.6B.Q4_K_M.gguf","size":484220032,"quant":"Q4_K_M","sha256":"dc4503da5d7cc7254055a86cd90e1a8c9d16c6ac71eb3a32b34bf48a1f4e0999"},{"name":"Qwen3-0.6B.Q5_K_M.gguf","size":551378048,"quant":"Q5_K_M","sha256":"33930feb827837f7e2827829b6af8a583ea7eb4c4f44096922e5c7f371645fc7"},{"name":"Qwen3-0.6B.Q6_K.gguf","size":622733440,"quant":"Q6_K","sha256":"d70ab3ca7859985909d2cd0ca23716c121a18c79d57b3827752d71c9ecda1e3e"},{"name":"Qwen3-0.6B.fp16.gguf","size":1509347456,"sha256":"7a6a16669751ca72ddc1c829b38f4520cfaa5da3b363adb11d79bbb1ccdea66d"}],"totalBytes":4364289920,"suggestedFile":"Qwen3-0.6B.Q4_K_M.gguf","requirements":{"ramGb":2,"diskBytes":484220032,"note":"estimate: suggested file size × 1.15 + 0.5 GB; a real measurement comes with lsh models"},"runWith":["llama.cpp","ollama"],"command":"lsh models install hf:MaziyarPanahi/Qwen3-0.6B-GGUF"}}