Store › model › LLM
Spark-X2.5-4B-GGUF
by XHToken · source Hugging Face · updated 2026-09-20
apache-2.02.4 GB~4 GB RAMsource aliveunlabeled
Spark-X2.5 is a compact, general-purpose language model for conversation, writing, translation, reasoning, coding, tool use, and agentic workflows. It uses a hybrid attention architecture, supports a…
Add to LogiShell Open in the web IDE
The button opens LogiShell with this card; nothing installs from a link by itself. Inside the app the install goes through lsh models install hf:XHToken/Spark-X2.5-4B-GGUF and its progress lives in the Resource Center.
Source and license
- Source: https://huggingface.co/XHToken/Spark-X2.5-4B-GGUF
- License: apache-2.0
- Requirements: about 4 GB of RAM, 2.4 GB on disk (estimate: suggested file size × 1.15 + 0.5 GB; a real measurement comes with lsh models). Runs with llama.cpp, ollama.
- Tags:
ggufllama.cppollamalm-studiosparkx2_5unslothpocketpal-aitext-generationenzhendpoints_compatibleconversational
Numbers
- 466,393 downloads on Hugging Face
- 256 likes
- license apache-2.0
- 2.4 GB for Spark-X2.5-4B-Q4_K_M.gguf
- 130,156 stars on ggml-org/llama.cpp
- 2,528 open issues and PRs
- last release v0.5.0 on 2026-09-23
- 327,838 npm downloads a week for node-llama-cpp
- latest node-llama-cpp@3.22.1
Numbers as of 2026-10-02 20:53 UTC, from the source API.
Summary
Summary not ready yet: the numbers are here, the text is not. It is written by the collector through the LogiShell model facade when a provider key is present.
Reviews
No reviews yet. Reviews are written inside LogiShell: open this card in the app.
Files
| file | quant | size |
|---|---|---|
Spark-X2.5-4B-Q4_K_M.gguf | Q4_K_M | 2.4 GB |
Spark-X2.5-4B-Q8_0.gguf | Q8_0 | 4.1 GB |
Spark-X2.5-4B.gguf | 7.7 GB |
From the source README
Spark-X2.5-4B-GGUF
Spark-X2.5 is a compact, general-purpose language model for conversation, writing, translation, reasoning, coding, tool use, and agentic workflows. It uses a hybrid attention architecture, supports a native context length of up to 1M tokens, and covers more than 200 languages. For its architecture, training methods, benchmark results, fine-tuning, and citation, see the Spark-X2.5-4B.
Local Deployment
llama.cpp
> [!NOTE]
> Native `spark2_5` support requires llama.cpp `b10828` or later.
Run the model in the CLI:
```bash
llama cli -m /path/to/Spark-X2.5-4B-GGUF
```
Launch the OpenAI-compatible API server:
```bash
llama serve -m /path/to/Spark-X2.5-4B-GGUF
```
Ollama
Download Ollama
> [!NOTE]
> Native `spark2_5` support requires Ollama `v0.34.1` or later.
ollama run SparkLLM/Spark-X2.5-4B
LM Studio
Card id model:hf:XHToken/Spark-X2.5-4B-GGUF · collected 2026-10-02 20:53 UTC · JSON