Tools
Well said @RamaswmySridhar! I believe the cost saving is actually much bigger, more like 4-5x. E.g., we have just reduced our GLM 5.2 cached…
Well said @RamaswmySridhar! I believe the cost saving is actually much bigger, more like 4-5x. E.g., we have just reduced our GLM 5.2 cached token price by 2X to return efficiency gain to our users. A
Well said @RamaswmySridhar! I believe the cost saving is actually much bigger, more like 4-5x. E.g., we have just reduced our GLM 5.2 cached token price by 2X to return efficiency gain to our users. Also just launched GLM 5.2 post training with zero-KLD. Frontier quality compounds business value. More to come. But here's the punchline. Normalized to 90% cache hit rate: GLM-5.2 (Fireworks): 1.12/session Opus-4.7 (Anthropic): 2.14/session GLM is ~48% cheaper.
Source: Fireworks AI (X) | 2026-06-26