Tested Muse Glimmer locally on coding with OpenCode & agentic work
DGX agentRan the model with quants (Q4) by Unsloth with latest (build from master) llama.cpp server. It takes ~20GB ram running on M5 Pro with 48GB at about 17t/s. Didn't do any reasoning loops/overthinking. O