Open models to add in one click
Collected from Hugging Face, GitHub and npm: open LLMs in GGUF, speech models like Whisper, embedding models. Every card links to its source, repeats the license and the numbers, and adds a short summary. One button adds it to LogiShell; without the app it leads to the install and brings you back.
-
granite-speech-4.1-2b-gguf
model · Speech · by handy-computer
granite-speech-4.1-2b: transcribe.cpp GGUF
apache-2.01.5 GB~3 GB RAMsource alive
36K downloads · 0 likes · numbers as of 2026-10-02
-
diar_streaming_sortformer_4spk-v2
model · Speech · by nvidia
🚨 Announcement: New Model Version Available
cc-by-4.0140 MB~1 GB RAMsource alive
34K downloads · 145 likes · numbers as of 2026-10-02
-
Fun-ASR-MLT-Nano-2512-gguf
model · Speech · by handy-computer
Fun-ASR-MLT-Nano-2512: transcribe.cpp GGUF
other531 MB~2 GB RAMsource alive
33K downloads · 1 likes · numbers as of 2026-10-02
-
multitalker-parakeet-streaming-0.6b-v1-gguf
model · Speech · by handy-computer
multitalker-parakeet-streaming-0.6b-v1: transcribe.cpp GGUF
other589 MB~2 GB RAMsource alive
33K downloads · 2 likes · numbers as of 2026-10-02
-
mxbai-embed-xsmall-v1
model · Embeddings · by mixedbread-ai
<path fill="#e94c28" d="M922 537c-6.003 11.784-11.44 23.81-19.66 34.428-6.345 8.196-11.065 17.635-17.206 26.008-4.339 5.916-9.828 10.992-14.854 16.397-.776.835-1.993 1.279-2.71 2.147-9.439 11.437-22.…
apache-2.029 MB~1 GB RAMsource alive
26K downloads · 37 likes · numbers as of 2026-10-02
-
embeddinggemma-300m-gguf
model · Embeddings · by ente-ai
EmbeddingGemma 300M Q80 GGUF — Ente mirror for Ensu
gemma318 MB~1 GB RAMsource alive
22K downloads · 0 likes · numbers as of 2026-10-02
-
SheetSage2-GGUF
model · Embeddings · by audio-cpp
GGUF conversion of m-a-p/SheetSage2 for audio.cpp. SheetSage2 transcribes music into melody, chords, beats, key, structure, and editable scores. This is a converted checkpoint, not a new model or an…
cc-by-nc-4.02.5 GB~4 GB RAMsource alive
21K downloads · 8 likes · numbers as of 2026-10-02
-
nemo-nano-codec-22khz-1.89kbps-21.5fps
model · Embeddings · by nvidia
The NeMo NanoCodec is a neural audio codec that leverages finite scalar quantization and adversarial training with large speech language models to achieve state-of-the-art audio compression across di…
other75 MB~1 GB RAMsource alive
18K downloads · 20 likes · numbers as of 2026-10-02
-
embedder_collection
model · Embeddings · by kalle07
A practical introduction to embedding models and local RAG
license unknown610 MB~2 GB RAMsource alive
15K downloads · 30 likes · numbers as of 2026-10-02
-
embeddinggemma-300m-GGUF
model · Embeddings · by unsloth
Responsible Generative AI Toolkit EmbeddingGemma on Kaggle EmbeddingGemma on Vertex Model Garden
gemma265 MB~1 GB RAMsource alive
14K downloads · 88 likes · numbers as of 2026-10-02
-
embeddinggemma-300M-qat-q4_0-GGUF
model · Embeddings · by ggml-org
Alternatively, the llama-embedding command line tool can be used: sh llama-embedding -hf ggml-org/embeddinggemma-300M-qat-q40-GGUF --verbose-prompt -p "Hello embeddings"
gemma265 MB~1 GB RAMsource alive
14K downloads · 6 likes · numbers as of 2026-10-02
-
bge-m3-Q8_0-GGUF
model · Embeddings · by ggml-org
ggml-org/bge-m3-Q80-GGUF This model was converted to GGUF format from BAAI/bge-m3 using llama.cpp via the ggml.ai's GGUF-my-repo space. Refer to the original model card for more details on the model.
mit605 MB~2 GB RAMsource alive
12K downloads · 17 likes · numbers as of 2026-10-02
-
nomic-embed-text-v1-GGUF
model · Embeddings · by nomic-ai
Embedding text with nomic-embed-text requires task instruction prefixes at the beginning of each string.
apache-2.080 MB~1 GB RAMsource alive
9.1K downloads · 7 likes · numbers as of 2026-10-02
-
LFM2.5-Embedding-350M-GGUF
model · Embeddings · by LiquidAI
LFM2.5-Embedding-350M is a dense bi-encoder for fast multilingual retrieval. It produces a single vector per document — the smallest, fastest index — for reliable cross-lingual search across 11 langu…
other219 MB~1 GB RAMsource alive
7.4K downloads · 39 likes · numbers as of 2026-10-02
-
gemma4-12b-with-proj-ltx-2.5-GGUF
model · Embeddings · by elix3r
embedding model by elix3r, gguf files on Hugging Face.
other7.8 GB~10 GB RAMsource alive
6.3K downloads · 23 likes · numbers as of 2026-10-02
-
gte-small-gguf
model · Embeddings · by ChristianAzinn
General Text Embeddings (GTE) model. Towards General Text Embeddings with Multi-stage Contrastive Learning
mit28 MB~1 GB RAMsource alive
5.9K downloads · 5 likes · numbers as of 2026-10-02
-
Qwen3.5-9B-DFlash-GGUF
model · Embeddings · by Anbeeld
GGUF quantizations of z-lab DFlash draft model for Qwen 3.5 9B.
apache-2.0730 MB~2 GB RAMsource alive
4.7K downloads · 5 likes · numbers as of 2026-10-02
-
jina-embeddings-v5-text-nano-retrieval-GGUF
model · Embeddings · by jinaai
jina-embeddings-v5-text-nano-retrieval-GGUF
cc-by-nc-4.0150 MB~1 GB RAMsource alive
4.6K downloads · 4 likes · numbers as of 2026-10-02
-
Qwen3.6-35B-A3B-DFlash-GGUF
model · Embeddings · by Anbeeld
GGUF quantizations of z-lab DFlash draft model for Qwen 3.6 35B A3B.
apache-2.0225 MB~1 GB RAMsource alive
4.4K downloads · 13 likes · numbers as of 2026-10-02
-
Ollama_Quantised_LLaVA-NeXT-Video-7B-DPO
model · Embeddings · by ManishThota
embedding model by ManishThota, gguf files on Hugging Face.
license unknown6.7 GB~9 GB RAMsource alive
3.9K downloads · 7 likes · numbers as of 2026-10-02
-
E5-Mistral-7B-Instruct-Embedding-GGUF
model · Embeddings · by second-state
embedding model by second-state, gguf files on Hugging Face.
mit4.1 GB~6 GB RAMsource alive
3.8K downloads · 15 likes · numbers as of 2026-10-02
-
Qwen3.8-27B-DSpark-GGUF
model · Embeddings · by Anbeeld
GGUF quantizations of RadixArk DSpark draft model for Qwen3.8 27B.
other1.0 GB~2 GB RAMsource alive
3.4K downloads · 2 likes · numbers as of 2026-10-02
-
jina-embeddings-v5-text-nano-text-matching
model · Embeddings · by jinaai
jina-embeddings-v5-text: Task-Targeted Embedding Distillation
cc-by-nc-4.0150 MB~1 GB RAMsource alive
3.3K downloads · 7 likes · numbers as of 2026-10-02
-
gte-large-gguf
model · Embeddings · by ChristianAzinn
General Text Embeddings (GTE) model. Towards General Text Embeddings with Multi-stage Contrastive Learning
mit206 MB~1 GB RAMsource alive
2.9K downloads · 1 likes · numbers as of 2026-10-02