Store › model › Speech
whisper-large-v3-turbo-gguf
by handy-computer · source Hugging Face · updated 2026-09-15
apache-2.0511 MB~2 GB RAMsource aliveunlabeled
whisper-large-v3-turbo: transcribe.cpp GGUF
Add to LogiShell Open in the web IDE
The button opens LogiShell with this card; nothing installs from a link by itself. Inside the app the install goes through lsh models install hf:handy-computer/whisper-large-v3-turbo-gguf and its progress lives in the Resource Center.
Source and license
- Source: https://huggingface.co/handy-computer/whisper-large-v3-turbo-gguf
- License: apache-2.0
- Requirements: about 2 GB of RAM, 511 MB on disk (estimate: suggested file size × 1.15 + 0.5 GB; a real measurement comes with lsh models). Runs with whisper.cpp.
- Tags:
transcribe.cppggufasrspeech-to-textwhisperopenaiautomatic-speech-recognitionafamarasazbabebgbn
Numbers
- 341,294 downloads on Hugging Face
- 4 likes
- license apache-2.0
- 0.5 GB for whisper-large-v3-turbo-Q4_K_M.gguf
- 15,195 npm downloads a week for nodejs-whisper
- latest nodejs-whisper@0.3.1
Numbers as of 2026-10-02 20:59 UTC, from the source API.
Summary
Summary not ready yet: the numbers are here, the text is not. It is written by the collector through the LogiShell model facade when a provider key is present.
Reviews
No reviews yet. Reviews are written inside LogiShell: open this card in the app.
Files
| file | quant | size |
|---|---|---|
whisper-large-v3-turbo-F16.gguf | F16 | 1.5 GB |
whisper-large-v3-turbo-Q4_K_M.gguf | Q4_K_M | 511 MB |
whisper-large-v3-turbo-Q5_K_M.gguf | Q5_K_M | 591 MB |
whisper-large-v3-turbo-Q6_K.gguf | Q6_K | 660 MB |
whisper-large-v3-turbo-Q8_0.gguf | Q8_0 | 845 MB |
From the source README
whisper-large-v3-turbo: transcribe.cpp GGUF
GGUF conversions of openai/whisper-large-v3-turbo for use
with transcribe.cpp.
Ported from upstream commit
41f01f3,
pinned 2026-04-25.
Validated against the transformers reference at transcribe.cpp commit
5.6.1
on 2026-04-26.
OpenAI Whisper large-v3-turbo — converted to GGUF for transcribe.cpp. Multilingual transcription and language detection; unlike the full large-v3 model, this turbo variant does not support speech translation. The v3 family adds Cantonese (yue) and uses a 128-bin mel input. Encoder-decoder transformer; 30-second windows with chunked long-form decoding.
Downloads
| Quantization | Download | Size | WER (LibriSpeech test-clean) |
| --- | --- | ---: | ---: |
| F16 | whisper-large-v3-turbo-F16.gguf | 1.63 GB | 2.01% |
| Q8_0 | whisper-large-v3-turbo-Q8_0.gguf | 886 MB | 2.01% |
| Q6_K | whisper-large-v3-turbo-Q6_K.gguf | 693 MB | 2.01% |
| Q5_K_M | whisper-large-v3-turbo-Q5_K_M.gguf | 620 MB | 2.03% |
| Q4_K_M | whisper-large-v3-turbo-Q4_K_M.gguf | 536 MB | 2.04% |
WER on the full LibriSpeech test-clean split (2,620 u…
Source: https://huggingface.co/handy-computer/whisper-large-v3-turbo-gguf
Card id model:hf:handy-computer/whisper-large-v3-turbo-gguf · collected 2026-10-02 20:59 UTC · JSON