LogiShell store Open app

Store › model › Embeddings

jina-embeddings-v5-text-nano-clustering

by jinaai · source Hugging Face · updated 2026-04-15

cc-by-nc-4.0150 MB~1 GB RAMsource aliveunlabeled

jina-embeddings-v5-text: Task-Targeted Embedding Distillation

Add to LogiShell Open in the web IDE

The button opens LogiShell with this card; nothing installs from a link by itself. Inside the app the install goes through lsh models install hf:jinaai/jina-embeddings-v5-text-nano-clustering and its progress lives in the Resource Center.

Source and license

Numbers

Numbers as of 2026-10-02 20:59 UTC, from the source API.

Summary

Summary not ready yet: the numbers are here, the text is not. It is written by the collector through the LogiShell model facade when a provider key is present.

Reviews

No reviews yet. Reviews are written inside LogiShell: open this card in the app.

Files

filequantsize
model.safetensors404 MB
onnx/model.onnx89 KB
onnx/model_fp16.onnx90 KB
onnx/model_q4.onnx126 KB
onnx/model_q4f16.onnxF16127 KB
onnx/model_quantized.onnx128 KB
v5-nano-clustering-F16.ggufF16411 MB
v5-nano-clustering-IQ1_M.ggufIQ1_M97 MB
v5-nano-clustering-IQ1_S.ggufIQ1_S95 MB
v5-nano-clustering-IQ2_M.ggufIQ2_M108 MB
v5-nano-clustering-IQ2_XXS.ggufIQ2_XXS101 MB
v5-nano-clustering-IQ4_NL.ggufIQ4_NL145 MB
v5-nano-clustering-IQ4_XS.ggufIQ4_XS142 MB
v5-nano-clustering-Q2_K.ggufQ2_K124 MB
v5-nano-clustering-Q3_K_M.ggufQ3_K_M137 MB
v5-nano-clustering-Q4_K_M.ggufQ4_K_M150 MB
v5-nano-clustering-Q5_K_M.ggufQ5_K_M161 MB
v5-nano-clustering-Q5_K_S.ggufQ5_K_S159 MB
v5-nano-clustering-Q6_K.ggufQ6_K173 MB
v5-nano-clustering-Q8_0.ggufQ8_0222 MB

From the source README

jina-embeddings-v5-text: Task-Targeted Embedding Distillation

Elastic Inference Service | ArXiv | Release Note | Blog

Model Overview

`jina-embeddings-v5-text-nano-clustering` is a compact, high-performance text embedding model designed for clustering.

It is part of the jina-embeddings-v5-text model family, which also includes jina-embeddings-v5-text-small, for better performance at a bigger size.

Trained using a novel approach that combines distillation with task-specific contrastive losses, `jina-embeddings-v5-text-nano-clustering` outperforms existing state-of-the-art models of similar size across diverse embedding benchmarks.
| Feature | Value |
| --- | --- |
| Parameters | 239M |
| Supported Tasks | `clustering` |
| Max Sequence Length | 8192 |
| Embedding Dimension | 768 |
| Matryoshka Dimensions | 32, 64, 128, 256, 512, 768 |
| Pooling Strategy | Last-token pooling |
| Base Model | jinaai/jina-embeddings-v5-text-nano |

Training and Evaluation

For training details and evaluation results, see our technical report.

Usage

Requirements

The following Python packages are required:

  • `transformers>=5.1.0`
  • `torch>=2.8.0`
  • `peft>=0.15.2`
  • `vllm==0.15.1`

Card id model:hf:jinaai/jina-embeddings-v5-text-nano-clustering · collected 2026-10-02 20:59 UTC · JSON