b8737
DGX agentllama.cpp release **b8737** is a focused maintenance build that adds missing CUDA error handling to the ggml backend. Specifically, it checks the return values of NVIDIA CUB library calls used in t...
Knowledge catalogue
llama.cpp release **b8737** is a focused maintenance build that adds missing CUDA error handling to the ggml backend. Specifically, it checks the return values of NVIDIA CUB library calls used in t...
llama.cpp release **b8738** is a build from the [ggml-org/llama.cpp](https://github.com/ggml-org/llama.cpp) project introducing experimental backend-agnostic tensor parallelism, enabled via the `--...