{"v":1,"id":"model:hf:unsloth/gpt-oss-20b-GGUF","slug":"model-unsloth-gpt-oss-20b-gguf","kind":"model","category":"llm","title":"gpt-oss-20b-GGUF","summary":"See our collection for all versions of gpt-oss including GGUF, 4-bit & 16-bit formats. Learn to run gpt-oss correctly - Read our Guide.","source":{"provider":"hf","ref":"unsloth/gpt-oss-20b-GGUF","url":"https://huggingface.co/unsloth/gpt-oss-20b-GGUF","rev":"d449b42d93e1c2c7bda5312f5c25c8fb91dfa9b4","fetchedAt":"2026-10-02T20:53:17.247Z","etag":"W/\"5e4a-OLTyQL1qQEAMs8ACnvsYCTuhYJ0\""},"author":{"name":"unsloth","url":"https://huggingface.co/unsloth"},"license":{"spdx":"apache-2.0","raw":"apache-2.0","open":true},"metrics":{"downloads":490693,"downloadsWeek":327838,"likes":849,"stars":130156,"openIssues":2528,"lastRelease":{"tag":"v0.5.0","at":"2026-09-23T20:50:06Z"},"pushedAt":"2026-10-02T20:17:08Z","takenAt":"2026-10-02T20:53:17.247Z"},"tags":["transformers","gguf","gpt_oss","text-generation","openai","unsloth","endpoints_compatible","conversational"],"pipeline":"text-generation","links":{"github":"ggml-org/llama.cpp","npm":"node-llama-cpp"},"updatedAt":"2025-12-19T01:39:27.000Z","collectedAt":"2026-10-02T20:53:17.247Z","review":{"numbers":["490,693 downloads on Hugging Face","849 likes","license apache-2.0","11 GB for gpt-oss-20b-Q4_K_M.gguf","130,156 stars on ggml-org/llama.cpp","2,528 open issues and PRs","last release v0.5.0 on 2026-09-23","327,838 npm downloads a week for node-llama-cpp","latest node-llama-cpp@3.22.1"],"log":null},"trust":"unlabeled","health":{"status":"alive","checkedAt":"2026-10-02T20:53:17.247Z","http":200},"description":"# Read our How to [Run gpt-oss Guide here!](https://docs.unsloth.ai/models/gpt-oss-how-to-run-and-fine-tune)\n\n  \n    See our collection for all versions of gpt-oss including GGUF, 4-bit & 16-bit formats.\n  \n  \n    Learn to run gpt-oss correctly - Read our Guide.\n  \n\n   See Unsloth Dynamic 2.0 GGUFs for our quantization benchmarks.\n  \n  \n    \n      \n    \n    \n      \n    \n    \n      \n    \n  \n✨ Read our gpt-oss Guide here!\n\n- Fine-tune gpt-oss-20b for free using our [Google Colab notebook](https://colab.research.google.com/github/unslothai/notebooks/blob/main/nb/gpt-oss-(20B)-Fine-tuning.ipynb)\n- Read our Blog about gpt-oss support: [unsloth.ai/blog/gpt-oss](https://unsloth.ai/blog/gpt-oss)\n- View the rest of our notebooks in our [docs here](https://docs.unsloth.ai/get-started/unsloth-notebooks).\n- Thank you to the [llama.cpp](https://github.com/ggml-org/llama.cpp) team for their work on supporting this model. We wouldn't be able to release quants without them!\n\nThe F32 quant is MXFP4 upcasted to BF16 for every single layer and is unquantized.\n\n# gpt-oss-20b Details\n\n  \n\n  Try gpt-oss ·\n  Guides ·\n  System card ·\n  OpenAI blog\n\nWelcome to the gpt-oss series, [OpenAI’s open-weight models](https://openai.com/open-models) designed for powerful reasoning, agentic tasks, and versatile developer use cases.\n\nWe’re releasing two flavors of the open models:\n- `gpt-oss-120b` — for production, general purpose, high reasoning use cases that fits into a single H100 GPU (117B parameters with 5.1B active parameters)\n- `gpt-oss-20b` — for lower latency, and local or specialized use cases (21B parameters with 3.6B active parameters)\n\nBoth models were trained on our [harmony response format](https://github.com/openai/harmony) and should only be used with the harmony format as it will not work correctly otherwise.\n\n> [!NOTE]\n> This model card is dedicated to the smaller `gpt-oss-20b` model. Check out [`gpt-oss-120b`](https://huggi…\n\nSource: https://huggingface.co/unsloth/gpt-oss-20b-GGUF","install":{"kind":"model","hfId":"unsloth/gpt-oss-20b-GGUF","gated":false,"format":"gguf","files":[{"name":"gpt-oss-20b-F16.gguf","size":13792639168,"quant":"F16","sha256":"4e4f9cd88d6456e4f389e7262eca4a8d565211e2b22ece9ca7a8556168ff3c66"},{"name":"gpt-oss-20b-Q2_K.gguf","size":11468317888,"quant":"Q2_K","sha256":"129b8e3865517cbba55bbb7d4f0cc2444a014e73a619eafbac247ef6573a19ee"},{"name":"gpt-oss-20b-Q2_K_L.gguf","size":11757884608,"quant":"Q2_K_L","sha256":"14de8eca18d52ae76449e0869bfb215686b212716787356b5bdd20793f82c6dc"},{"name":"gpt-oss-20b-Q3_K_M.gguf","size":11506103488,"quant":"Q3_K_M","sha256":"bc02b57e05cfef36b1f6f4a952666a76b26838f8cf431412cb0cd19f41cf8040"},{"name":"gpt-oss-20b-Q3_K_S.gguf","size":11463894208,"quant":"Q3_K_S","sha256":"8deacd2c5ed46d1a62368c52e44abdb55e32a2aa0e5acea86d18d8258339f17d"},{"name":"gpt-oss-20b-Q4_0.gguf","size":11501495488,"quant":"Q4_0","sha256":"8f5890ffd832518b1843b8e1b39125c70fb2c104a6e88613d70f45efa6aaba0e"},{"name":"gpt-oss-20b-Q4_1.gguf","size":11577504448,"quant":"Q4_1","sha256":"de7e287610b8d792090cc3b001fe91957d3738197e53af83f763810c06d519e0"},{"name":"gpt-oss-20b-Q4_K_M.gguf","size":11624759488,"quant":"Q4_K_M","sha256":"c27536640e410032865dc68781d80a08b98f8db5e93575919af8ccc0568aeb4f"},{"name":"gpt-oss-20b-Q4_K_S.gguf","size":11618492608,"quant":"Q4_K_S","sha256":"3c5483e8749f4865f1aae2c36b796e7b8c43ec02c5a74d663a82a5fe916b2298"},{"name":"gpt-oss-20b-Q5_K_M.gguf","size":11717357248,"quant":"Q5_K_M","sha256":"9c3814533c5b4c84d42b5dce4376bbdfd7227e990b8733a3a1c4f741355b3e75"},{"name":"gpt-oss-20b-Q5_K_S.gguf","size":11711827648,"quant":"Q5_K_S","sha256":"b9a5c5312dcf63af3ab5a2a2aed6e3bc86d51d452bce364ff44f573a8b089141"},{"name":"gpt-oss-20b-Q6_K.gguf","size":12041000128,"quant":"Q6_K","sha256":"f2ead52a7842a556ea10d5edd91a88e5af974856d2485ecaad3aade679832935"},{"name":"gpt-oss-20b-Q8_0.gguf","size":12109567168,"quant":"Q8_0","sha256":"bcd455d4034ec02f71a875b46cb17df44a97911d7258291973be4d21f98329f3"},{"name":"gpt-oss-20b-UD-Q4_K_XL.gguf","size":11872347328,"quant":"Q4_K_XL","sha256":"10fe673de12c20b74b8d670a9fdf0fd36b43b0a86ffc04daeb175c0a2b98c4f9"},{"name":"gpt-oss-20b-UD-Q6_K_XL.gguf","size":12041000128,"quant":"Q6_K_XL","sha256":"f2ead52a7842a556ea10d5edd91a88e5af974856d2485ecaad3aade679832935"},{"name":"gpt-oss-20b-UD-Q8_K_XL.gguf","size":13195442368,"quant":"Q8_K_XL","sha256":"b97fa9f3efa642ef3abd599a170f0b26612e5179a4ef5c65fb8852fb974ec239"}],"totalBytes":190999633408,"suggestedFile":"gpt-oss-20b-Q4_K_M.gguf","requirements":{"ramGb":13,"diskBytes":11624759488,"note":"estimate: suggested file size × 1.15 + 0.5 GB; a real measurement comes with lsh models"},"runWith":["llama.cpp","ollama"],"command":"lsh models install hf:unsloth/gpt-oss-20b-GGUF"}}