{"v":1,"id":"model:hf:jinaai/jina-embeddings-v5-text-small-text-matching","slug":"model-jinaai-jina-embeddings-v5-text-small-text-matching","kind":"model","category":"embedding","title":"jina-embeddings-v5-text-small-text-matching","summary":"jina-embeddings-v5-text-small-text-matching: Text-Matching-Targeted Embedding Distillation","source":{"provider":"hf","ref":"jinaai/jina-embeddings-v5-text-small-text-matching","url":"https://huggingface.co/jinaai/jina-embeddings-v5-text-small-text-matching","rev":"69890c317b81d4e1f412e66ebe68b6ebf9534d24","fetchedAt":"2026-10-02T21:00:32.562Z","etag":"W/\"3cdd-xkKG3xv1cR2XDXwo0wHJmJvCxoU\""},"author":{"name":"jinaai","url":"https://huggingface.co/jinaai"},"license":{"spdx":"cc-by-nc-4.0","raw":"cc-by-nc-4.0","open":false,"note":"restricted license: terms at the source apply"},"metrics":{"downloads":116726,"downloadsWeek":327838,"likes":13,"takenAt":"2026-10-02T21:00:32.562Z"},"tags":["llama.cpp","onnx","safetensors","gguf","qwen3","embedding","llama-cpp","jina-embeddings-v5","feature-extraction","mteb","vllm","sentence-transformers","sentence-similarity","multilingual","text-embeddings-inference","conversational"],"pipeline":"sentence-similarity","links":{"github":"ggml-org/llama.cpp","npm":"node-llama-cpp"},"updatedAt":"2026-04-15T09:40:09.000Z","collectedAt":"2026-10-02T21:00:32.562Z","review":{"numbers":["116,726 downloads on Hugging Face","13 likes","license cc-by-nc-4.0","0.4 GB for v5-small-text-matching-Q4_K_M.gguf","327,838 npm downloads a week for node-llama-cpp","latest node-llama-cpp@3.22.1"],"log":null},"trust":"unlabeled","health":{"status":"alive","checkedAt":"2026-10-02T21:00:32.562Z","http":200},"description":"### **jina-embeddings-v5-text-small-text-matching**: Text-Matching-Targeted Embedding Distillation\n\n[Elastic Inference Service](https://www.elastic.co/docs/explore-analyze/elastic-inference/eis) | [ArXiv](https://arxiv.org/abs/2602.15547) | [Release Note](https://jina.ai/news/jina-embeddings-v5-text-distilling-4b-quality-into-sub-1b-multilingual-embeddings) | [Blog](https://www.elastic.co/search-labs/blog/jina-embeddings-v5-text)\n\n### Model Overview\n\n`jina-embeddings-v5-text-small-text-matching` is a compact, high-performance text embedding model designed for text-matching.\n\nIt is part of the **jina-embeddings-v5-text** model family, which also includes [jina-embeddings-v5-text-nano](https://huggingface.co/jinaai/jina-embeddings-v5-text-nano), a smaller model for more resource-constrained use cases.\n\nTrained using a novel approach that combines distillation with task-specific contrastive losses, `jina-embeddings-v5-text-small-text-matching` outperforms existing state-of-the-art models of similar size across diverse embedding benchmarks.\n| Feature | Value |\n| --- | --- |\n| Parameters | 677M |\n| Supported Tasks | `text-matching`|\n| Max Sequence Length | 32768 |\n| Embedding Dimension | 1024 |\n| Matryoshka Dimensions | 32, 64, 128, 256, 512, 768, 1024 |\n| Pooling Strategy | Last-token pooling |\n| Base Model | jinaai/jina-embeddings-v5-text-small |\n\n### Training and Evaluation\n\nFor training details and evaluation results, see our [technical report](https://arxiv.org/abs/2602.15547).\n\n### Usage\n\n  Requirements\n\nThe following Python packages are required:\n\n- `transformers>=5.1.0`\n- `torch>=2.8.0`\n- `peft>=0.15.2`\n- `vllm>=0.15.1`\n\n### Optional / Recommended\n- **flash-attention**: Installing [flash-attention](https://github.com/Dao-AILab/flash-attention) is recommended for improved inference speed and efficiency, but not mandatory.\n- **sentence-transformers**: If you want to use the model vi…\n\nSource: https://huggingface.co/jinaai/jina-embeddings-v5-text-small-text-matching","install":{"kind":"model","hfId":"jinaai/jina-embeddings-v5-text-small-text-matching","gated":false,"format":"gguf","files":[{"name":"model.safetensors","size":1192133232,"sha256":"030f08ea0e2be8a58eb549a9daa90e9ed3b5db3f562df2625eee26b3ef1c5baf"},{"name":"onnx/model.onnx","size":1266078,"sha256":"463fb3fec2adf9c47a946626dc2a0884e43091f95466083fb1754950000a7c1a"},{"name":"v5-small-text-matching-F16.gguf","size":1198182464,"quant":"F16","sha256":"f42b976b9e508dc143ab2dfd984e130548397276227a08340b227a2cc6c18878"},{"name":"v5-small-text-matching-IQ1_M.gguf","size":216052096,"quant":"IQ1_M","sha256":"f62ede1ac67cab8af78b9b83b5fcc7a8b2fe961f7ac32abfd1ebdeb42526afc7"},{"name":"v5-small-text-matching-IQ1_S.gguf","size":208015744,"quant":"IQ1_S","sha256":"0a392783f032e421ea95681b1d0979c5fb777dfb4801588e10f79cebac3dea65"},{"name":"v5-small-text-matching-IQ2_M.gguf","size":264909184,"quant":"IQ2_M","sha256":"2f34f8fa5a1354a8a47d622f6e9d806bdb7dd4a86b42252a59d246bcae77bed5"},{"name":"v5-small-text-matching-IQ2_XXS.gguf","size":229446016,"quant":"IQ2_XXS","sha256":"97623291ebc7cc507cf8ea927e6350ae09d273a3cd931ede2969dd2967dd6cfe"},{"name":"v5-small-text-matching-IQ4_NL.gguf","size":381566336,"quant":"IQ4_NL","sha256":"e1180dad9cc07ab6e85f811a1feeb17eee11db6b385297443371dfaac1580f4e"},{"name":"v5-small-text-matching-IQ4_XS.gguf","size":367803776,"quant":"IQ4_XS","sha256":"625760498c3628ed0b866b46433f993dae722e6a9fabcbb15c69e32d0a0adb97"},{"name":"v5-small-text-matching-Q2_K.gguf","size":296238464,"quant":"Q2_K","sha256":"c9db454105a23f999c85d79f7ff8c2bc9759140e920aed86d4d899c3dc09a099"},{"name":"v5-small-text-matching-Q3_K_M.gguf","size":347127168,"quant":"Q3_K_M","sha256":"989515bb3b6b81b88c9d5646c286521d1c25f051584f37bee04c4f9a220205f9"},{"name":"v5-small-text-matching-Q4_K_M.gguf","size":396705152,"quant":"Q4_K_M","sha256":"738555454772b436632c6bad5891aeaa38d414bd7d7185107caeb3b2d8f2d860"},{"name":"v5-small-text-matching-Q5_K_M.gguf","size":444415360,"quant":"Q5_K_M","sha256":"ad5025091c572bb450dd08a6afb5bcd80cfd239d12d8f5f1c3718ba19e9d9053"},{"name":"v5-small-text-matching-Q5_K_S.gguf","size":436616576,"quant":"Q5_K_S","sha256":"7bd9dbdd4bb694cee889195e77444b0932645c41a4e0c770614e5e8ac98a0b67"},{"name":"v5-small-text-matching-Q6_K.gguf","size":495107456,"quant":"Q6_K","sha256":"2f485b41ceed799765b141871c5e72997e9e3706e8666579072870f554d0c76d"},{"name":"v5-small-text-matching-Q8_0.gguf","size":639447424,"quant":"Q8_0","sha256":"0cab4e0788c3b9bdb42d9063374509762bba31316a38736e06ebc753793df334"}],"totalBytes":7115032526,"suggestedFile":"v5-small-text-matching-Q4_K_M.gguf","requirements":{"ramGb":1,"diskBytes":396705152,"note":"estimate: suggested file size × 1.15 + 0.5 GB; a real measurement comes with lsh models"},"runWith":["llama.cpp"],"command":"lsh models install hf:jinaai/jina-embeddings-v5-text-small-text-matching"}}