Applications
2026: the year of explosive AI innovation. @FireworksAI_HQ Co-founder & CEO Lin Qiao speaks on how the massive expansion of customized infer…
2026: the year of explosive AI innovation. @FireworksAI_HQ Co-founder & CEO Lin Qiao speaks on how the massive expansion of customized inference and models is creating speed of light production to sca
2026: the year of explosive AI innovation. @FireworksAI_HQ Co-founder & CEO Lin Qiao speaks on how the massive expansion of customized inference and models is creating speed of light production to scale timelines. Media
Related
- At 30 trillion tokens a day, we're the largest inference provider outside the frontier labs' own APIs. We're hiring. Come build the infrastr…
- ECHO: Elastic Speculative Decoding with Sparse Gating for High-Concurrency Scenarios
- GLM 5.1 is live on Fireworks! SOTA for agents and coding: →Plans and executes multi-hour workflows without falling apart →Planning, executin…
- From Tokens to Layers: Redefining Stall-Free Scheduling for MoE Serving with Layered Prefill
- We are excited to have @FireworksAI_HQ as a day 0 launch partner for Kimi K2.6! Their inference and fine-tuning platform is fast, reliable, …
Source: Fireworks AI (X) | 2026-04-28