{"v":1,"id":"model:hf:FINAL-Bench/POCKET-35B-GGUF","slug":"model-final-bench-pocket-35b-gguf","kind":"model","category":"llm","title":"POCKET-35B-GGUF","summary":"🆕 POCKET-Qwen3.8-Flash-Next — a 180B model running on a laptop with 8 GB VRAM + 32 GB RAM · 4.17 tok/s measured. >","source":{"provider":"hf","ref":"FINAL-Bench/POCKET-35B-GGUF","url":"https://huggingface.co/FINAL-Bench/POCKET-35B-GGUF","rev":"c9e42c51357fdd5f25af7cd7dc0e8bf59c81af4d","fetchedAt":"2026-10-02T20:59:47.536Z","etag":"W/\"2c60-7oLChP/CcR6mOOjR9OT+0enKXN0\""},"author":{"name":"FINAL-Bench","url":"https://huggingface.co/FINAL-Bench"},"license":{"spdx":"apache-2.0","raw":"apache-2.0","open":true},"metrics":{"downloads":669265,"downloadsWeek":327838,"likes":81,"takenAt":"2026-10-02T20:59:47.536Z"},"tags":["llama.cpp","gguf","conversational","on-device","mobile","iphone","android","cpu","local-llm","edge","mixture-of-experts","moe","quantized","pocket","vidraft","qwen3_5_moe","darwin","text-generation","endpoints_compatible","imatrix"],"pipeline":"text-generation","links":{"github":"ggml-org/llama.cpp","npm":"node-llama-cpp"},"updatedAt":"2026-09-25T12:53:11.000Z","collectedAt":"2026-10-02T20:59:47.536Z","review":{"numbers":["669,265 downloads on Hugging Face","81 likes","license apache-2.0","20 GB for POCKET-35B-Q4_K_M.gguf","327,838 npm downloads a week for node-llama-cpp","latest node-llama-cpp@3.22.1"],"log":null},"trust":"unlabeled","health":{"status":"alive","checkedAt":"2026-10-02T20:59:47.536Z","http":200},"description":"> ### 🆕 **[POCKET-Qwen3.8-Flash-Next](https://huggingface.co/FINAL-Bench/POCKET-Qwen3.8-Flash-Next-GGUF)** — a **180B** model running on a **laptop with 8 GB VRAM + 32 GB RAM** · **4.17 tok/s measured**.\n> [](https://huggingface.co/FINAL-Bench/POCKET-Qwen3.8-Flash-Next-GGUF) []() []() []()\n\n> ### 🆕 **[POCKET-Zimage-CPU](https://huggingface.co/FINAL-Bench/POCKET-Zimage-CPU)** — photoreal **images** in **46 s** on a **CPU only**. No GPU, no CUDA, no Python.\n> [](https://huggingface.co/FINAL-Bench/POCKET-Zimage-CPU) [](https://huggingface.co/spaces/FINAL-Bench/POCKET-Zimage-CPU) []()\n\n> ### 📚 Collections\n> **▶ [POCKET Models](https://huggingface.co/collections/FINAL-Bench/pocket-models-6a618ee5d23eafb7e185a5c6)** — this family (on-device, no GPU)\n> [Darwin Family](https://huggingface.co/collections/FINAL-Bench/darwin-family-699987b1f652864af0122193) · [Aether Foundation](https://huggingface.co/collections/FINAL-Bench/aether-foundation-model-6a5c7f2fa1a4165c0414e53a) · [VKAE Accelerated](https://huggingface.co/collections/FINAL-Bench/vkae-accelerated-6a47231d7e7999dd8227675a)\n\n# POCKET-35B-GGUF\n\n### A **35B** model that runs on your **PC with no GPU** — and on your phone. Just stock `llama.cpp`. **No fork**, no CUDA, no cloud.\n\n> 🚀 **Try it live, no install →** [](https://huggingface.co/spaces/FINAL-Bench/POCKET-35B-CPU) [](https://huggingface.co/spaces/FINAL-Bench/POCKET-26B-CPU) — both answering on a **CPU-only** box (no GPU). POCKET-26B is Gemma4-based.\n\n[](https://www.apache.org/licenses/LICENSE-2.0) [](https://github.com/ggml-org/llama.cpp) []() []()\n\n**Pick your build →** [](https://huggingface.co/FINAL-Bench/POCKET-35B-GGUF) [](https://huggingface.co/FINAL-Bench/POCKET-26B-GGUF) [](https://huggingface.co/FINAL-Bench/POCKET-KR-GGUF) [-0f6e56)](https://huggingface.co/FINAL-Bench/POCKET-KR-MLX) [](https://huggingface.co/FINAL-Bench/POCKET-EN-GGUF) [](https://huggingface.co/FINAL-Bench/POCKET-Qwen3.8-Fl…\n\nSource: https://huggingface.co/FINAL-Bench/POCKET-35B-GGUF","install":{"kind":"model","hfId":"FINAL-Bench/POCKET-35B-GGUF","gated":false,"format":"gguf","files":[{"name":"POCKET-35B-IQ1_M.gguf","size":8239208384,"quant":"IQ1_M","sha256":"c56c77d158786f6dfb084e8cf3467456668e400cc7632c366cc0444b9f8565a1"},{"name":"POCKET-35B-Q2_K.gguf","size":12939593664,"quant":"Q2_K","sha256":"2567ed710fb5b5bdf0df99ecfa241ab1a2b35edc618d2830b91d468e3a3ce191"},{"name":"POCKET-35B-Q3_K_M.gguf","size":16764763840,"quant":"Q3_K_M","sha256":"9ab4184f0f5af0cc1cf8eee7e553664b692c3f837a8868c4a16137bb3a3c97a2"},{"name":"POCKET-35B-Q4_K_M.gguf","size":21166757568,"quant":"Q4_K_M","sha256":"6f479f637c8fb932df39b9cfabdc454568eee48c0e7c0584e1815a27558e8ffe"}],"totalBytes":59110323456,"suggestedFile":"POCKET-35B-Q4_K_M.gguf","requirements":{"ramGb":24,"diskBytes":21166757568,"note":"estimate: suggested file size × 1.15 + 0.5 GB; a real measurement comes with lsh models"},"runWith":["llama.cpp","ollama"],"command":"lsh models install hf:FINAL-Bench/POCKET-35B-GGUF"}}