LogiShell store Open app

Store › model › Embeddings

mxbai-embed-large-v1-gguf

by ChristianAzinn · source Hugging Face · updated 2024-04-07

apache-2.0206 MB~1 GB RAMsource aliveunlabeled

This is our base sentence embedding model. It was trained using AnglE loss on our high-quality large scale data. It achieves SOTA performance on BERT-large scale. Find out more in our blog post.

Add to LogiShell Open in the web IDE

The button opens LogiShell with this card; nothing installs from a link by itself. Inside the app the install goes through lsh models install hf:ChristianAzinn/mxbai-embed-large-v1-gguf and its progress lives in the Resource Center.

Source and license

Numbers

Numbers as of 2026-10-02 21:01 UTC, from the source API.

Summary

Summary not ready yet: the numbers are here, the text is not. It is written by the collector through the LogiShell model facade when a provider key is present.

Reviews

No reviews yet. Reviews are written inside LogiShell: open this card in the app.

Files

filequantsize
mxbai-embed-large-v1.Q2_K.ggufQ2_K138 MB
mxbai-embed-large-v1.Q3_K_L.ggufQ3_K_L189 MB
mxbai-embed-large-v1.Q3_K_M.ggufQ3_K_M173 MB
mxbai-embed-large-v1.Q3_K_S.ggufQ3_K_S152 MB
mxbai-embed-large-v1.Q4_0.ggufQ4_0190 MB
mxbai-embed-large-v1.Q4_K_M.ggufQ4_K_M206 MB
mxbai-embed-large-v1.Q4_K_S.ggufQ4_K_S194 MB
mxbai-embed-large-v1.Q5_0.ggufQ5_0226 MB
mxbai-embed-large-v1.Q5_K_M.ggufQ5_K_M234 MB
mxbai-embed-large-v1.Q5_K_S.ggufQ5_K_S226 MB
mxbai-embed-large-v1.Q6_K.ggufQ6_K265 MB
mxbai-embed-large-v1.Q8_0.ggufQ8_0342 MB
mxbai-embed-large-v1_fp16.gguf639 MB
mxbai-embed-large-v1_fp32.gguf1.2 GB

From the source README

mxbai-embed-large-v1-gguf

Model creator: MixedBread AI

Original model: mxbai-embed-large-v1

Original Description

This is our base sentence embedding model. It was trained using AnglE loss on our high-quality large scale data. It achieves SOTA performance on BERT-large scale. Find out more in our blog post.

Description

This repo contains GGUF format files for the mxbai-embed-large-v1 embedding model.

These files were converted and quantized with llama.cpp PR 5500, commit 34aa045de, on a consumer RTX 4090.

This model supports up to 512 tokens of context.

Compatibility

These files are compatible with llama.cpp as of commit 4524290e8, as well as LM Studio as of version 0.2.19.

# Meta-information
## Explanation of quantisation methods

Card id model:hf:ChristianAzinn/mxbai-embed-large-v1-gguf · collected 2026-10-02 21:01 UTC · JSON