Store › model › Speech
VibeVoice-ASR-Streaming-1.5B-GGUF
by christopherthompson81 · source Hugging Face · updated 2026-09-20
mit3.1 GB~5 GB RAMsource aliveunlabeled
VibeVoice ASR Streaming 1.5B GGUF for audio.cpp
Add to LogiShell Open in the web IDE
The button opens LogiShell with this card; nothing installs from a link by itself. Inside the app the install goes through lsh models install hf:christopherthompson81/VibeVoice-ASR-Streaming-1.5B-GGUF and its progress lives in the Resource Center.
Source and license
- Source: https://huggingface.co/christopherthompson81/VibeVoice-ASR-Streaming-1.5B-GGUF
- License: mit
- Requirements: about 5 GB of RAM, 3.1 GB on disk (estimate: suggested file size × 1.15 + 0.5 GB; a real measurement comes with lsh models). Runs with whisper.cpp.
- Tags:
ggufaudio.cppasrstreamingspeaker-attributed-transcriptionautomatic-speech-recognitionenzhesptdejakofrruit
Numbers
- 39,182 downloads on Hugging Face
- 0 likes
- license mit
- 3.1 GB for vibevoice-asr-streaming-1.5b-q8_0.gguf
- 54,249 stars on ggml-org/whisper.cpp
- 352 open issues and PRs
- last release v1.9.5 on 2026-10-06
- 15,664 npm downloads a week for nodejs-whisper
- latest nodejs-whisper@0.3.1
Numbers as of 2026-10-09 19:01 UTC, from the source API.
Summary
Summary not ready yet: the numbers are here, the text is not. It is written by the collector through the LogiShell model facade when a provider key is present.
Reviews
No reviews yet. Reviews are written inside LogiShell: open this card in the app.
Files
| file | quant | size |
|---|---|---|
vibevoice-asr-streaming-1.5b-bf16.gguf | BF16 | 5.3 GB |
vibevoice-asr-streaming-1.5b-q4_k.gguf | Q4_K | 2.0 GB |
vibevoice-asr-streaming-1.5b-q8_0.gguf | Q8_0 | 3.1 GB |
From the source README
VibeVoice ASR Streaming 1.5B GGUF for audio.cpp
This repository contains audio.cpp-native GGUF builds of
`microsoft/VibeVoice-ASR-Streaming-1.5B`.
The 1.5B checkpoint runs through the same audio.cpp loader as the 7B with no
code changes: the layer count, hidden size, and head counts are read from the
checkpoint's own `config.json`, and the tensor names are identical.
Use with audio.cpp
Install the recommended Q8_0 package through the audio.cpp model manager:
python3 tools/model_manager_v2.py install vibevoice_asr_streaming_1_5b_q8_0
Run offline ASR:
build/debug/bin/audiocpp_cli \
--task asr \
--family vibevoice_asr_streaming \
--model models/VibeVoice-ASR-Streaming-1.5B-GGUF/vibevoice-asr-streaming-1.5b-q8_0.gguf \
--backend cuda \
--threads 8 \
--audio input.wav \
--text-out transcript.txt \
--metrics \
--log
Run the server with the model loaded:
{
"host": "127.0.0.1",
"port": 8080,
"backend": "cuda",
"threads": 8,
"models": [
{
"id": "vibevoice-streaming-1.5b",
"family": "vibevoice_asr_streaming",
"path": "models/VibeVoice-ASR-Streaming-1.5B-GGUF/vibevoice-asr-streaming-1.5b-q8_0.gguf",
"task": "asr",
"mode": "streaming"
}
]
}
Then start the server:
build/debug/bin/audiocpp_server --config server.json --log
Card id model:hf:christopherthompson81/VibeVoice-ASR-Streaming-1.5B-GGUF · collected 2026-10-09 19:01 UTC · JSON