Store › model › LLM
Spark-X2.5-4B-GGUF
by abenzerps · source Hugging Face · updated 2026-09-07
apache-2.02.4 GB~4 GB RAMsource aliveunlabeled
Compatibility: These GGUF files require llama.cpp b10828 or later, which includes official support for the Spark-X2.5 (spark25) architecture. Applications with a bundled runtime must use an equivalen…
Add to LogiShell Open in the web IDE
The button opens LogiShell with this card; nothing installs from a link by itself. Inside the app the install goes through lsh models install hf:abenzerps/Spark-X2.5-4B-GGUF and its progress lives in the Resource Center.
Source and license
- Source: https://huggingface.co/abenzerps/Spark-X2.5-4B-GGUF
- License: apache-2.0
- Requirements: about 4 GB of RAM, 2.4 GB on disk (estimate: suggested file size × 1.15 + 0.5 GB; a real measurement comes with lsh models). Runs with llama.cpp, ollama.
- Tags:
ggufllama.cppspark-x2.5long-context1m-contexttext-generationconversationalendpoints_compatible
Numbers
- 318,878 downloads on Hugging Face
- 24 likes
- license apache-2.0
- 2.4 GB for Spark-X2.5-4B-Q4_K_M.gguf
- 130,649 stars on ggml-org/llama.cpp
- 2,491 open issues and PRs
- last release v0.6.0 on 2026-10-05
- 297,971 npm downloads a week for node-llama-cpp
- latest node-llama-cpp@3.22.1
Numbers as of 2026-10-09 19:01 UTC, from the source API.
Summary
Summary not ready yet: the numbers are here, the text is not. It is written by the collector through the LogiShell model facade when a provider key is present.
Reviews
No reviews yet. Reviews are written inside LogiShell: open this card in the app.
Files
| file | quant | size |
|---|---|---|
Spark-X2.5-4B-Q4_0.gguf | Q4_0 | 2.2 GB |
Spark-X2.5-4B-Q4_K_M.gguf | Q4_K_M | 2.4 GB |
Spark-X2.5-4B-Q5_K_M.gguf | Q5_K_M | 2.8 GB |
Spark-X2.5-4B-Q6_K.gguf | Q6_K | 3.1 GB |
Spark-X2.5-4B-Q8_0.gguf | Q8_0 | 4.1 GB |
From the source README
> [!IMPORTANT]
> Compatibility: These GGUF files require `llama.cpp` b10828 or later, which includes official support for the Spark-X2.5 (`spark2_5`) architecture. Applications with a bundled runtime must use an equivalent or newer build. llama.cpp support
Spark-X2.5-4B GGUF
GGUF quantizations of XHToken/Spark-X2.5-4B, a 4B general-purpose language model for reasoning, coding, tool use, and agentic workflows. Native context: 1,048,576 tokens (1M).
Benchmarks
*Benchmark results reported by XHToken for Spark-X2.5-4B in thinking mode.*
GGUF files
| Quantization | File | Size |
| --- | --- | ---: |
| Q4_0 | Spark-X2.5-4B-Q4_0.gguf | 2.41 GB |
| Q4_K_M | Spark-X2.5-4B-Q4_K_M.gguf | 2.60 GB |
| Q5_K_M | Spark-X2.5-4B-Q5_K_M.gguf | 2.98 GB |
| Q6_K | Spark-X2.5-4B-Q6_K.gguf | 3.38 GB |
| Q8_0 | Spark-X2.5-4B-Q8_0.gguf | 4.38 GB |
Includes the upstream chat_template.jinja. Checksums: SHA256SUMS.txt.
Usage
llama-cli -m Spark-X2.5-4B-Q4_K_M.gguf -c 131072 -cnv
Source
- Model: XHToken/Spark-X2.5-4B
- Revision: `ea14618d20e76b5b093d3ee20a5b9d733bb12410`
- License: Apache-2.0
Card id model:hf:abenzerps/Spark-X2.5-4B-GGUF · collected 2026-10-09 19:01 UTC · JSON