Local Ai
b9466
B9466 is a build release in the llama.cpp project that includes fixes and improvements to speculative decoding functionality, specifically addressing n_outputs_max issues and extracting helper functio
B9466 is a build release in the llama.cpp project that includes fixes and improvements to speculative decoding functionality, specifically addressing n_outputs_max issues and extracting helper functions for speculative inference optimization. Llama.cpp enables LLM inference with minimal setup and state-of-the-art performance on a wide range of hardware locally and in the cloud.
Source: llama.cpp Releases | 2026-06-02