Store › model › Embeddings
mxbai-embed-large-v1-gguf
by ChristianAzinn · source Hugging Face · updated 2024-04-07
apache-2.0206 MB~1 GB RAMsource aliveunlabeled
This is our base sentence embedding model. It was trained using AnglE loss on our high-quality large scale data. It achieves SOTA performance on BERT-large scale. Find out more in our blog post.
Add to LogiShell Open in the web IDE
The button opens LogiShell with this card; nothing installs from a link by itself. Inside the app the install goes through lsh models install hf:ChristianAzinn/mxbai-embed-large-v1-gguf and its progress lives in the Resource Center.
Source and license
- Source: https://huggingface.co/ChristianAzinn/mxbai-embed-large-v1-gguf
- License: apache-2.0
- Requirements: about 1 GB of RAM, 206 MB on disk (estimate: suggested file size × 1.15 + 0.5 GB; a real measurement comes with lsh models). Runs with llama.cpp.
- Tags:
sentence-transformersggufmtebtransformerstransformers.jsfeature-extractionendeploy:azure
Numbers
- 238,659 downloads on Hugging Face
- 8 likes
- license apache-2.0
- 0.2 GB for mxbai-embed-large-v1.Q4_K_M.gguf
- 4,402,710 npm downloads a week for @huggingface/transformers
- latest @huggingface/transformers@4.3.0
Numbers as of 2026-10-02 21:01 UTC, from the source API.
Summary
Summary not ready yet: the numbers are here, the text is not. It is written by the collector through the LogiShell model facade when a provider key is present.
Reviews
No reviews yet. Reviews are written inside LogiShell: open this card in the app.
Files
| file | quant | size |
|---|---|---|
mxbai-embed-large-v1.Q2_K.gguf | Q2_K | 138 MB |
mxbai-embed-large-v1.Q3_K_L.gguf | Q3_K_L | 189 MB |
mxbai-embed-large-v1.Q3_K_M.gguf | Q3_K_M | 173 MB |
mxbai-embed-large-v1.Q3_K_S.gguf | Q3_K_S | 152 MB |
mxbai-embed-large-v1.Q4_0.gguf | Q4_0 | 190 MB |
mxbai-embed-large-v1.Q4_K_M.gguf | Q4_K_M | 206 MB |
mxbai-embed-large-v1.Q4_K_S.gguf | Q4_K_S | 194 MB |
mxbai-embed-large-v1.Q5_0.gguf | Q5_0 | 226 MB |
mxbai-embed-large-v1.Q5_K_M.gguf | Q5_K_M | 234 MB |
mxbai-embed-large-v1.Q5_K_S.gguf | Q5_K_S | 226 MB |
mxbai-embed-large-v1.Q6_K.gguf | Q6_K | 265 MB |
mxbai-embed-large-v1.Q8_0.gguf | Q8_0 | 342 MB |
mxbai-embed-large-v1_fp16.gguf | 639 MB | |
mxbai-embed-large-v1_fp32.gguf | 1.2 GB |
From the source README
mxbai-embed-large-v1-gguf
Model creator: MixedBread AI
Original model: mxbai-embed-large-v1
Original Description
This is our base sentence embedding model. It was trained using AnglE loss on our high-quality large scale data. It achieves SOTA performance on BERT-large scale. Find out more in our blog post.
Description
This repo contains GGUF format files for the mxbai-embed-large-v1 embedding model.
These files were converted and quantized with llama.cpp PR 5500, commit 34aa045de, on a consumer RTX 4090.
This model supports up to 512 tokens of context.
Compatibility
These files are compatible with llama.cpp as of commit 4524290e8, as well as LM Studio as of version 0.2.19.
# Meta-information
## Explanation of quantisation methods
Card id model:hf:ChristianAzinn/mxbai-embed-large-v1-gguf · collected 2026-10-02 21:01 UTC · JSON