Store › model › Embeddings
sentence-bert-swedish-cased
by KBLab · source Hugging Face · updated 2025-11-07
apache-2.0239 MB~1 GB RAMsource aliveunlabeled
This is a sentence-transformers model: It maps Swedish sentences & paragraphs to a 768 dimensional dense vector space and can be used for tasks like clustering or semantic search. This model is a bil…
Add to LogiShell Open in the web IDE
The button opens LogiShell with this card; nothing installs from a link by itself. Inside the app the install goes through lsh models install hf:KBLab/sentence-bert-swedish-cased and its progress lives in the Resource Center.
Source and license
- Source: https://huggingface.co/KBLab/sentence-bert-swedish-cased
- License: apache-2.0
- Requirements: about 1 GB of RAM, 239 MB on disk (estimate: suggested file size × 1.15 + 0.5 GB; a real measurement comes with lsh models). Runs with llama.cpp.
- Tags:
sentence-transformerspytorchsafetensorsggufbertfeature-extractionsentence-similaritytransformerssvtext-embeddings-inferenceendpoints_compatibledeploy:azure
Numbers
- 69,299 downloads on Hugging Face
- 36 likes
- license apache-2.0
- 0.2 GB for Sentence-Bert-Swedish-Cased-124M-BF16.gguf
- 327,838 npm downloads a week for node-llama-cpp
- latest node-llama-cpp@3.22.1
Numbers as of 2026-10-02 21:00 UTC, from the source API.
Summary
Summary not ready yet: the numbers are here, the text is not. It is written by the collector through the LogiShell model facade when a provider key is present.
Reviews
No reviews yet. Reviews are written inside LogiShell: open this card in the app.
Files
| file | quant | size |
|---|---|---|
Sentence-Bert-Swedish-Cased-124M-BF16.gguf | BF16 | 239 MB |
model.safetensors | 476 MB | |
pytorch_model.bin | 476 MB |
From the source README
KBLab/sentence-bert-swedish-cased
This is a sentence-transformers model: It maps Swedish sentences & paragraphs to a 768 dimensional dense vector space and can be used for tasks like clustering or semantic search. This model is a bilingual Swedish-English model trained according to instructions in the paper Making Monolingual Sentence Embeddings Multilingual using Knowledge Distillation and the documentation accompanying its companion python package. We have used the strongest available pretrained English Bi-Encoder (all-mpnet-base-v2) as a teacher model, and the pretrained Swedish KB-BERT as the student model.
A more detailed description of the model can be found in an article we published on the KBLab blog here and for the updated model here.
Update: We have released updated versions of the model since the initial release. The original model described in the blog post is v1.0. The current version is v2.0. The newer versions are trained on longer paragraphs, and have a longer max sequence length. v2.0 is trained with a stronger teacher model and is the current default.
| Model version | Teacher Model | Max Sequence Length |
|---------------|---------|----------|
| v1.0 | paraphrase-mpnet-base-v2 | 256 |
| v1.1 | paraphrase-mpnet-base-v2 | 384 |
| v2.0 | [all-mpnet-base-v2](https://huggingface.co/sentence-tran…
Source: https://huggingface.co/KBLab/sentence-bert-swedish-cased
Card id model:hf:KBLab/sentence-bert-swedish-cased · collected 2026-10-02 21:00 UTC · JSON