Observations on Muse-Glimmer reasoning traces being noticeably different from qwen / gemma models and questions for you guys
DGX agentJust downloaded the model, UD-Q5_K_XL quant, asked it to generate a long story to test out reasoning and speed with dflash (super fast btw, ~ 90 to 160 tok/s on a 5090 depending on task) and was surpr