LogiShell store Open app

Store › model › LLM

Spark-X2.5-4B-GGUF

by XHToken · source Hugging Face · updated 2026-09-20

apache-2.02.4 GB~4 GB RAMsource aliveunlabeled

Spark-X2.5 is a compact, general-purpose language model for conversation, writing, translation, reasoning, coding, tool use, and agentic workflows. It uses a hybrid attention architecture, supports a…

Add to LogiShell Open in the web IDE

The button opens LogiShell with this card; nothing installs from a link by itself. Inside the app the install goes through lsh models install hf:XHToken/Spark-X2.5-4B-GGUF and its progress lives in the Resource Center.

Source and license

Numbers

Numbers as of 2026-10-02 20:53 UTC, from the source API.

Summary

Summary not ready yet: the numbers are here, the text is not. It is written by the collector through the LogiShell model facade when a provider key is present.

Reviews

No reviews yet. Reviews are written inside LogiShell: open this card in the app.

Files

filequantsize
Spark-X2.5-4B-Q4_K_M.ggufQ4_K_M2.4 GB
Spark-X2.5-4B-Q8_0.ggufQ8_04.1 GB
Spark-X2.5-4B.gguf7.7 GB

From the source README

Spark-X2.5-4B-GGUF

Spark-X2.5 is a compact, general-purpose language model for conversation, writing, translation, reasoning, coding, tool use, and agentic workflows. It uses a hybrid attention architecture, supports a native context length of up to 1M tokens, and covers more than 200 languages. For its architecture, training methods, benchmark results, fine-tuning, and citation, see the Spark-X2.5-4B.

Local Deployment

llama.cpp

> [!NOTE]
> Native `spark2_5` support requires llama.cpp `b10828` or later.

Run the model in the CLI:
```bash
llama cli -m /path/to/Spark-X2.5-4B-GGUF
```

Launch the OpenAI-compatible API server:
```bash
llama serve -m /path/to/Spark-X2.5-4B-GGUF
```

Ollama

Download Ollama

> [!NOTE]
> Native `spark2_5` support requires Ollama `v0.34.1` or later.

ollama run SparkLLM/Spark-X2.5-4B

LM Studio

Card id model:hf:XHToken/Spark-X2.5-4B-GGUF · collected 2026-10-02 20:53 UTC · JSON