Local Ai
b9488
Release b9488 of llama.cpp was published on June 3, 2026 , and includes support for Qwen3 SSM architectures with additions like LLM_KV_ATTENTION_RECURRENT_LAYERS . The release also fixes a bug in comm
Release b9488 of llama.cpp was published on June 3, 2026 , and includes support for Qwen3 SSM architectures with additions like LLM_KV_ATTENTION_RECURRENT_LAYERS . The release also fixes a bug in common_prompt_batch_decode affecting session state storage and restoration, ensuring all tokens are properly stored in session_tokens .
Source: llama.cpp Releases | 2026-06-03