Store › model › Embeddings
Qwen3-Embedding-0.6B-GGUF
by mradermacher · source Hugging Face · updated 2026-04-23
apache-2.0378 MB~1 GB RAMsource aliveunlabeled
static quants of https://huggingface.co/Qwen/Qwen3-Embedding-0.6B
Add to LogiShell Open in the web IDE
The button opens LogiShell with this card; nothing installs from a link by itself. Inside the app the install goes through lsh models install hf:mradermacher/Qwen3-Embedding-0.6B-GGUF and its progress lives in the Resource Center.
Source and license
- Source: https://huggingface.co/mradermacher/Qwen3-Embedding-0.6B-GGUF
- License: apache-2.0
- Requirements: about 1 GB of RAM, 378 MB on disk (estimate: suggested file size × 1.15 + 0.5 GB; a real measurement comes with lsh models). Runs with llama.cpp.
- Tags:
transformersggufsentence-transformerssentence-similarityfeature-extractiontext-embeddings-inferenceenendpoints_compatibleconversational
Numbers
- 1,827 downloads on Hugging Face
- 2 likes
- license apache-2.0
- 0.4 GB for Qwen3-Embedding-0.6B.Q4_K_M.gguf
- 27,659 stars on QwenLM/Qwen3
- 68 open issues and PRs
- 297,971 npm downloads a week for node-llama-cpp
- latest node-llama-cpp@3.22.1
Numbers as of 2026-10-09 19:01 UTC, from the source API.
Summary
Summary not ready yet: the numbers are here, the text is not. It is written by the collector through the LogiShell model facade when a provider key is present.
Reviews
No reviews yet. Reviews are written inside LogiShell: open this card in the app.
Files
| file | quant | size |
|---|---|---|
Qwen3-Embedding-0.6B.IQ4_XS.gguf | IQ4_XS | 352 MB |
Qwen3-Embedding-0.6B.Q2_K.gguf | Q2_K | 282 MB |
Qwen3-Embedding-0.6B.Q3_K_L.gguf | Q3_K_L | 351 MB |
Qwen3-Embedding-0.6B.Q3_K_M.gguf | Q3_K_M | 331 MB |
Qwen3-Embedding-0.6B.Q3_K_S.gguf | Q3_K_S | 308 MB |
Qwen3-Embedding-0.6B.Q4_K_M.gguf | Q4_K_M | 378 MB |
Qwen3-Embedding-0.6B.Q4_K_S.gguf | Q4_K_S | 365 MB |
Qwen3-Embedding-0.6B.Q5_K_M.gguf | Q5_K_M | 424 MB |
Qwen3-Embedding-0.6B.Q5_K_S.gguf | Q5_K_S | 416 MB |
Qwen3-Embedding-0.6B.Q6_K.gguf | Q6_K | 472 MB |
Qwen3-Embedding-0.6B.Q8_0.gguf | Q8_0 | 610 MB |
Qwen3-Embedding-0.6B.f16.gguf | F16 | 1.1 GB |
From the source README
About
static quants of https://huggingface.co/Qwen/Qwen3-Embedding-0.6B
*For a convenient overview and download list, visit our model page for this model.*
weighted/imatrix quants are available at https://huggingface.co/mradermacher/Qwen3-Embedding-0.6B-i1-GGUF
## Usage
If you are unsure how to use GGUF files, refer to one of TheBloke's
READMEs for
more details, including on how to concatenate multi-part files.
Provided Quants
(sorted by size, not necessarily quality. IQ-quants are often preferable over similar sized non-IQ quants)
| Link | Type | Size/GB | Notes |
|:-----|:-----|--------:|:------|
| GGUF | Q2_K | 0.4 | |
| GGUF | Q3_K_S | 0.4 | |
| GGUF | Q3_K_M | 0.4 | lower quality |
| GGUF | Q3_K_L | 0.5 | |
| GGUF | IQ4_XS | 0.5 | |
| GGUF | Q4_K_S | 0.5 | fast, recommended |
| GGUF | Q4_K_M | 0.5 | fast, recommended |
| GGUF | Q5_K_S | 0.5 | |
| [GGUF](https://huggingface.co/mradermac…
Source: https://huggingface.co/mradermacher/Qwen3-Embedding-0.6B-GGUF
Card id model:hf:mradermacher/Qwen3-Embedding-0.6B-GGUF · collected 2026-10-09 19:01 UTC · JSON