Store › model › Speech
canary-1b-v2-gguf
by handy-computer · source Hugging Face · updated 2026-09-15
cc-by-4.0701 MB~2 GB RAMsource aliveunlabeled
GGUF conversions of nvidia/canary-1b-v2 for use with transcribe.cpp.
Add to LogiShell Open in the web IDE
The button opens LogiShell with this card; nothing installs from a link by itself. Inside the app the install goes through lsh models install hf:handy-computer/canary-1b-v2-gguf and its progress lives in the Resource Center.
Source and license
- Source: https://huggingface.co/handy-computer/canary-1b-v2-gguf
- License: cc-by-4.0
- Requirements: about 2 GB of RAM, 701 MB on disk (estimate: suggested file size × 1.15 + 0.5 GB; a real measurement comes with lsh models). Runs with whisper.cpp.
- Tags:
transcribe.cppggufasrspeech-to-textcanarymultitask-aedencoder-decodermultilingualtranslationautomatic-speech-recognitionbghrcsdanlen
Numbers
- 90,002 downloads on Hugging Face
- 1 likes
- license cc-by-4.0
- 0.7 GB for canary-1b-v2-Q4_K_M.gguf
Numbers as of 2026-10-02 20:59 UTC, from the source API.
Summary
Summary not ready yet: the numbers are here, the text is not. It is written by the collector through the LogiShell model facade when a provider key is present.
Reviews
No reviews yet. Reviews are written inside LogiShell: open this card in the app.
Files
| file | quant | size |
|---|---|---|
canary-1b-v2-F16.gguf | F16 | 1.8 GB |
canary-1b-v2-F32.gguf | F32 | 3.7 GB |
canary-1b-v2-Q4_K_M.gguf | Q4_K_M | 701 MB |
canary-1b-v2-Q5_K_M.gguf | Q5_K_M | 798 MB |
canary-1b-v2-Q6_K.gguf | Q6_K | 889 MB |
canary-1b-v2-Q8_0.gguf | Q8_0 | 1.1 GB |
From the source README
canary-1b-v2: transcribe.cpp GGUF
GGUF conversions of nvidia/canary-1b-v2 for use
with transcribe.cpp.
Ported from upstream commit
87bc526,
pinned 2026-05-08.
Validated against the NeMo reference at transcribe.cpp commit
db53eda
on 2026-05-08.
Offline multilingual speech-to-text and translation across 25 European
languages. A 978M-parameter multitask AED with a 32-layer FastConformer
encoder and an 8-layer Transformer decoder. Supports automatic speech
recognition for any of the 25 supported languages, plus translation
between supported language pairs (per the upstream model card). Takes
a 16 kHz mono WAV and produces a transcript. Not a streaming model;
word and segment timestamps from the upstream model are not exposed in
the v1 port.
Downloads
| Quantization | Download | Size | WER (LibriSpeech test-clean) |
| --- | --- | ---: | ---: |
| F32 | canary-1b-v2-F32.gguf | 3.92 GB | 1.92% |
| F16 | canary-1b-v2-F16.gguf | 1.97 GB | 1.92% |
| Q8_0 | canary-1b-v2-Q8_0.gguf | 1.14 GB | 1.91% |
| Q6_K | canary-1b-v2-Q6_K.gguf | 932 MB | 1.94% |
| Q5_K_M | canary-1b-v2-Q5_K_M.gguf | 837 MB | 1.93% |
| Q4_K_M | [canary-1b-v2-Q4_K_M.gguf](https://huggingface.co/handy-computer/canary-1b-v2-gguf/resolve/main/canary-1b…
Source: https://huggingface.co/handy-computer/canary-1b-v2-gguf
Card id model:hf:handy-computer/canary-1b-v2-gguf · collected 2026-10-02 20:59 UTC · JSON