LogiShell store Open app

Store › model › Embeddings

gte-small-gguf

by ChristianAzinn · source Hugging Face · updated 2024-04-07

mit28 MB~1 GB RAMsource aliveunlabeled

General Text Embeddings (GTE) model. Towards General Text Embeddings with Multi-stage Contrastive Learning

Add to LogiShell Open in the web IDE

The button opens LogiShell with this card; nothing installs from a link by itself. Inside the app the install goes through lsh models install hf:ChristianAzinn/gte-small-gguf and its progress lives in the Resource Center.

Source and license

Numbers

Numbers as of 2026-10-02 20:59 UTC, from the source API.

Summary

Summary not ready yet: the numbers are here, the text is not. It is written by the collector through the LogiShell model facade when a provider key is present.

Reviews

No reviews yet. Reviews are written inside LogiShell: open this card in the app.

Files

filequantsize
gte-small.Q2_K.ggufQ2_K24 MB
gte-small.Q3_K_L.ggufQ3_K_L26 MB
gte-small.Q3_K_M.ggufQ3_K_M25 MB
gte-small.Q3_K_S.ggufQ3_K_S24 MB
gte-small.Q4_0.ggufQ4_025 MB
gte-small.Q4_K_M.ggufQ4_K_M28 MB
gte-small.Q4_K_S.ggufQ4_K_S27 MB
gte-small.Q5_0.ggufQ5_028 MB
gte-small.Q5_K_M.ggufQ5_K_M29 MB
gte-small.Q5_K_S.ggufQ5_K_S28 MB
gte-small.Q6_K.ggufQ6_K33 MB
gte-small.Q8_0.ggufQ8_035 MB
gte-small_fp16.gguf64 MB
gte-small_fp32.gguf127 MB

From the source README

gte-small-gguf

Model creator: thenlper

Original model: gte-small

Original Description

General Text Embeddings (GTE) model. Towards General Text Embeddings with Multi-stage Contrastive Learning

The GTE models are trained by Alibaba DAMO Academy. They are mainly based on the BERT framework and currently offer three different sizes of models, including GTE-large, GTE-base, and GTE-small. The GTE models are trained on a large-scale corpus of relevance text pairs, covering a wide range of domains and scenarios. This enables the GTE models to be applied to various downstream tasks of text embeddings, including information retrieval, semantic textual similarity, text reranking, etc.

Description

This repo contains GGUF format files for the gte-small embedding model.

These files were converted and quantized with llama.cpp PR 5500, commit 34aa045de, on a consumer RTX 4090.

This model supports up to 512 tokens of context.

Compatibility

These files are compatible with llama.cpp as of commit 4524290e8, as well as LM Studio as of version 0.2.19.

Card id model:hf:ChristianAzinn/gte-small-gguf · collected 2026-10-02 20:59 UTC · JSON