Applications
At 30 trillion tokens a day, we're the largest inference provider outside the frontier labs' own APIs. We're hiring. Come build the infrastr…
At 30 trillion tokens a day, we're the largest inference provider outside the frontier labs' own APIs. We're hiring. Come build the infrastructure behind the top AI workloads: https://fireworks.ai/car
At 30 trillion tokens a day, we're the largest inference provider outside the frontier labs' own APIs. We're hiring. Come build the infrastructure behind the top AI workloads: https://fireworks.ai/careers 📸 @Wing VC's 2026 Enterprise Tech 30 list (thanks @WingVC @ericnewcomer)
Related
- GLM 5.1 is live on Fireworks! SOTA for agents and coding: →Plans and executes multi-hour workflows without falling apart →Planning, executin…
- There's no doubt that the world can consume tokens as fast as they're produced, even in the most maximalist infrastructure buildup scenarios…
- From Tokens to Layers: Redefining Stall-Free Scheduling for MoE Serving with Layered Prefill
- Excited to share our work on production-ready W4A8 inference, now integrated in vLLM! By combining 4-bit weights (low memory) with 8-bit act…
Source: Fireworks AI (X) | 2026-04-23