Model Releases
New Model: Spark-X2.5-4B, Spark-X2.5-1.7B
I was browsing HF for small LLMs and run into this model. It does not seem to be a fine tune - the model has its own architecture. https://huggingface.co/XHToken/Spark-X2.5-1.7B https://huggingface.co
I was browsing HF for small LLMs and run into this model. It does not seem to be a fine tune - the model has its own architecture. https://huggingface.co/XHToken/Spark-X2.5-1.7B https://huggingface.co/XHToken/Spark-X2.5-4B There are 4B/1.7B versions - the benchmark is quite interesting (4B is neck and neck with Qwen 3.5 9B). Currently does not run out of the box on llama.cpp - pending this PR: https://github.com/ggml-org/llama.cpp/pull/27868 They have a custom fork of llama.cpp that works. Anyone has tried this? Update: GGUFs (require custom fork for now): https://huggingface.co/XHToken/Spark-X2.5-1.7B-GGUF https://huggingface.co/XHToken/Spark-X2.5-4B-GGUF submitted by /u/insraq [link] [comments]
Related
- A.X-K2 released
- Smol king nanbeige 4.2 now with dspark!
- Agents-A1-4B (Qwen3.7-4B ???) : Scaling the Horizon, Not the Parameters
Source: r/LocalLLaMA | 2026-09-01