LogiShell store Open app

Store › model › LLM

s1-mini-GGUF

by superwhisper · source Hugging Face · updated 2026-08-28

other462 MB~2 GB RAMsource aliveunlabeled

GGUF builds of superwhisper/s1-mini, release v1, for llama.cpp, Ollama, LM Studio, and anything else built on llama.cpp. You can use it in your own dictation app too, just check the license first.

Add to LogiShell Open in the web IDE

The button opens LogiShell with this card; nothing installs from a link by itself. Inside the app the install goes through lsh models install hf:superwhisper/s1-mini-GGUF and its progress lives in the Resource Center.

Source and license

Numbers

Numbers as of 2026-10-09 19:01 UTC, from the source API.

Summary

Summary not ready yet: the numbers are here, the text is not. It is written by the collector through the LogiShell model facade when a provider key is present.

Reviews

No reviews yet. Reviews are written inside LogiShell: open this card in the app.

Files

filequantsize
s1-mini-f16.ggufF161.4 GB
s1-mini-q4_k_m.ggufQ4_K_M462 MB

From the source README

S1-mini-GGUF by Superwhisper

GGUF builds of superwhisper/s1-mini,
release v1, for llama.cpp, Ollama, LM Studio, and anything else built on
llama.cpp. You can use it in your own dictation app too, just check the
license first.

S1-mini is a 0.6B-parameter text normalizer for speech-to-text output. It
takes a raw ASR transcript and rewrites it as clean written text: fillers
removed, false starts and self-corrections resolved to the value the speaker
landed on, punctuation and capitalization applied, and spoken numbers, dates,
times, currency and email addresses rendered in written form.

At Q4_K_M it is a 462 MiB file that runs comfortably on a laptop CPU, and on a
held-out set of 7,519 English cases it reaches 94.8% token accuracy.

The model covers English only. It is not a chat model and will not follow
general instructions; it does one job, and you steer it with a control line at
the top of the input. Full documentation lives in the
BF16 repository.

Files

| File | Type | Size | Notes |
|---|---|---|---|
| `s1-mini-q4_k_m.gguf` | Q4_K_M | 462 MB | Recommended. The build the published accuracy was measured on. |
| `s1-mini-f16.gguf` | F16 | 1.4 GB | Unquantized conversion, the intermediate the Q4_K_M is produced from. |

Both files share the same skeleton: architecture `qwen3`, 311 tensors, 28
blocks, a 40,960-token context window, and an embedded chat template. The
Q4_K_M build keeps the most quantization-sensitive tensors at Q6_K (29 of 311)
and the bulk at Q4_K, while normalization parameters stay F32 in both builds.

> [!NOTE]
> The Hub sidebar reports 0.8B parameters for this repo. Qwen3-0.6B sets
> `tie_word_embeddings`, but ships `lm_head.weight…

Source: https://huggingface.co/superwhisper/s1-mini-GGUF

Card id model:hf:superwhisper/s1-mini-GGUF · collected 2026-10-09 19:01 UTC · JSON