AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

Search: “fireworks-ai--x”

GridTimelineEvolution
165 results
Applications

Loving this piece: bringing delight to your daily coffee, your daily workflows, and your AI bill. @gumloop is now running open-weight models…

DGX agent

Loving this piece: bringing delight to your daily coffee, your daily workflows, and your AI bill. @gumloop is now running open-weight models on Fireworks in production. 80% lower cost vs. closed-lab A

applicationsfireworks-ai--x
9 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Tools

Happy to announce that we’ve enlisted @FireworksAI_HQ as our primary provider for open-source models. In AI inference, quality, latency, rel…

DGX agent

Happy to announce that we’ve enlisted @FireworksAI_HQ as our primary provider for open-source models. In AI inference, quality, latency, reliability, and privacy vary a lot. Fireworks has earned our t

toolsfireworks-ai--x
8 Jul 2026
Tools

If you're tired of re-explaining the same context to your AI every day, go see what @PromptQL just launched!

DGX agent

If you're tired of re-explaining the same context to your AI every day, go see what @PromptQL just launched! We raised $136M to kill Slack. Introducing PromptQL: The first AI version of Slack. Here’s

toolsfireworks-ai--x
8 Jul 2026
Model Releases

Not another demo or benchmark. @ShoucongChen is a senior member of our technical staff. A real project, scoped at 1-month. Delivered in 4 da…

DGX agent

Not another demo or benchmark. @ShoucongChen is a senior member of our technical staff. A real project, scoped at 1-month. Delivered in 4 days with GLM5.2 Fast. The best devs deserve >400 t/sec. Take

model-releasesfireworks-ai--x
8 Jul 2026
Applications

This is what production RL infra looks like. ICYMI: RL training at scale separates into two distinct problems. - Tight collective comms for …

DGX agent

This is what production RL infra looks like. ICYMI: RL training at scale separates into two distinct problems. - Tight collective comms for the trainer. - Distributed async inference for rollout. Kudo

applicationsfireworks-ai--x
8 Jul 2026
Tools

Attending @RaiseSummit? Join our CEO @lqiao tomorrow @ 4:00 PM for a fireside chat with @MattEvantic of Evantic Capital on the Master Stage.…

DGX agent

Fireworks AI is hosting a fireside chat at Raise Summit featuring CEO Lqiao and Matt Evantic from Evantic Capital on the Master Stage at 4:00 PM. The event appears to be a discussion panel or intervie

toolsfireworks-ai--x
7 Jul 2026
Tutorials

What actually makes an AI application good or not? Thanks to @StackOverflow for hosting @the_bunny_chen on the Stack Overflow Podcast (even …

DGX agent

What actually makes an AI application good or not? Thanks to @StackOverflow for hosting @the_bunny_chen on the Stack Overflow Podcast (even though he is giving away all of our secrets!) Listen here: h

tutorialsfireworks-ai--x
7 Jul 2026
Tools

Today, we're helping kick off AMD AI Developer Hackathon: ACT II. We're giving $50 in Fireworks credits to participants in this event which …

DGX agent

Today, we're helping kick off AMD AI Developer Hackathon: ACT II. We're giving 50 in Fireworks credits to participants in this event which also includes: 20K in prizes. $150 in credits. Fully online.

toolsfireworks-ai--x
6 Jul 2026
Hardware

Don't forget, we'll be in Paris for @RaiseSummit. Will we see you there?

DGX agent

Don't forget, we'll be in Paris for @RaiseSummit. Will we see you there? Going to be in Paris for @RaiseSummit? Join us and @nvidia to hear from leaders at both companies on the future of AI inference

hardwarefireworks-ai--x
3 Jul 2026
Agents

Prompt engineering is costing you money. Learn how to fine-tune your models. Stuffing a lot of text into every single API call slows down yo…

DGX agent

Prompt engineering is costing you money. Learn how to fine-tune your models. Stuffing a lot of text into every single API call slows down your app because the model has to process all those tokens bef

agentsfireworks-ai--x
3 Jul 2026
Applications

Fear not, @FireworksAI_HQ is perfectly legal and serves the best models on 🇺🇸 servers for your AI-ndependence

DGX agent

Fear not, @FireworksAI_HQ is perfectly legal and serves the best models on 🇺🇸 servers for your AI-ndependence For info on firework enforcement, park closures and the like for this Fourth of July weeke

applicationsfireworks-ai--x
2 Jul 2026
Model Releases

do you know what you pay for in agentic workloads? cached tokens! session with 50+ tool calls -> prompt is billed 50 times all providers giv…

DGX agent

do you know what you pay for in agentic workloads? cached tokens! session with 50+ tool calls -> prompt is billed 50 times all providers give 1/5 cached discount for GLM-5.2 we at @FireworksAI_HQ drop

model-releasesfireworks-ai--x
1 Jul 2026
Tools

Fireworks Batch API: 50% cheaper than serverless. We obviously love things fast, but sometimes async at scale is all you need. With the refr…

DGX agent

Fireworks Batch API: 50% cheaper than serverless. We obviously love things fast, but sometimes async at scale is all you need. With the refreshed Batch API, you queue up a job and select whether you n

toolsfireworks-ai--x
1 Jul 2026
Applications

Frontier open models with enterprise governance means rebuilding workflows from scratch. Move faster with Fireworks on Foundry. GLM 5.2 is l…

DGX agent

Frontier open models with enterprise governance means rebuilding workflows from scratch. Move faster with Fireworks on Foundry. GLM 5.2 is live on Microsoft Foundry. With FireConnect enabled, devs can

applicationsfireworks-ai--x
1 Jul 2026
Agents

If you want frontier-level coding and agent performance but you don't want to pay closed-model prices, GLM 5.2 is the open model you've prob…

DGX agent

If you want frontier-level coding and agent performance but you don't want to pay closed-model prices, GLM 5.2 is the open model you've probably been hearing about. Here's why you should run it on Fir

agentsfireworks-ai--x
1 Jul 2026
Model Releases

This is exactly why we believe in customization. Quick context: Factory's original secret scanner was deterministic, so it either flagged th…

DGX agent

This is exactly why we believe in customization. Quick context: Factory's original secret scanner was deterministic, so it either flagged things that weren't actually secrets (false positives) or miss

model-releasesfireworks-ai--x
1 Jul 2026
Tutorials

For even higher speeds, reach out for a custom deployment! We’ve hit 446 tok/s on Artificial Analysis. Learn more → https://fireworks.ai/blo…

DGX agent

Fireworks AI announced achieving 446 tokens per second throughput speeds as measured by Artificial Analysis benchmarks, positioning this as their standard performance metric. The company offers custom

tutorialsfireworks-ai--x
30 Jun 2026
Applications

inference reliability has historically been a tax on devs that only large well-funded startups could afford: reserve GPUs in advance, sign a…

DGX agent

inference reliability has historically been a tax on devs that only large well-funded startups could afford: reserve GPUs in advance, sign a contract, guess your peak throughput requirements. everyone

applicationsfireworks-ai--x
30 Jun 2026
Tools

We heard your feedback. You want to go faster. Introducing GLM 5.2 Fast The same model and quality as GLM 5.2 standard, now at 140 tok/s Fli…

DGX agent

Fireworks AI announced GLM 5.2 Fast, an optimized version of their GLM 5.2 model that maintains the same quality and capabilities as the standard version while delivering significantly faster inferenc

toolsfireworks-ai--x
30 Jun 2026
Hardware

Going to be in Paris for @RaiseSummit? Join us and @nvidia to hear from leaders at both companies on the future of AI inference and infrastr…

DGX agent

Going to be in Paris for @RaiseSummit? Join us and @nvidia to hear from leaders at both companies on the future of AI inference and infrastructure. After that: cocktails and a DJ. Oh la la! See you th

hardwarefireworks-ai--x
29 Jun 2026
Agents

Research —> Product :) very excited to start rolling out our fine-tuned Trace Judge model built earlier this month with the great @Fireworks…

DGX agent

Research —> Product :) very excited to start rolling out our fine-tuned Trace Judge model built earlier this month with the great @FireworksAI_HQ team there’s a mountain of Agent Improvement gold sitt

agentsfireworks-ai--x
29 Jun 2026
Model Releases

you may have heard that glm-5.2 at 392 token/s is cool, how about 446 except… it’s all noise. Artificial Analysis picks median among 8 point…

DGX agent

you may have heard that glm-5.2 at 392 token/s is cool, how about 446 except… it’s all noise. Artificial Analysis picks median among 8 points/day so first point of the day can be way off looking at 3

model-releasesfireworks-ai--x
28 Jun 2026
Applications

DSpark from @deepseek_ai ingeniously integrates many speculative decoding ideas to achieve 1.5x to 5x higher throughput in a real production…

DGX agent

DSpark from @deepseek_ai ingeniously integrates many speculative decoding ideas to achieve 1.5x to 5x higher throughput in a real production system Let's understand it with 10 ideas, starting from the

applicationsfireworks-ai--x
27 Jun 2026
Applications

Model management is the real SDLC scaling bottleneck. @FactoryAI standardized on Fireworks to solve it: → 2–3x open-model growth → 5–15x mor…

DGX agent

Model management is the real SDLC scaling bottleneck. @FactoryAI standardized on Fireworks to solve it: → 2–3x open-model growth → 5–15x more work per dollar → day-0 access to every new open-weight mo

applicationsfireworks-ai--x
27 Jun 2026
Model Releases

Fireworks AI is now live on EvoSkill v1.3.0! You can now use @FireworksAI_HQ directly with EvoSkill to run fast inference on open models as …

DGX agent

Fireworks AI is now live on EvoSkill v1.3.0! You can now use @FireworksAI_HQ directly with EvoSkill to run fast inference on open models as both the evolution harness backend and the LLM scorer. Along

model-releasesfireworks-ai--x
26 Jun 2026
Applications

The big lesson from training @cursor_ai Composer 2: models exploit flaws in their training environment before learning what you actually wan…

DGX agent

The big lesson from training @cursor_ai Composer 2: models exploit flaws in their training environment before learning what you actually want. Real RL for coding agents means production-faithful envir

applicationsfireworks-ai--x
26 Jun 2026
Tools

We hosted the first RSI RL Environments hackathon with @hud_evals @ycombinator and it was a blast! We watched builders treat RL as a general…

DGX agent

We hosted the first RSI RL Environments hackathon with @hud_evals @ycombinator and it was a blast! We watched builders treat RL as a general purpose tool and reach for it across domains we had not ant

toolsfireworks-ai--x
26 Jun 2026
Tools

Well said @RamaswmySridhar! I believe the cost saving is actually much bigger, more like 4-5x. E.g., we have just reduced our GLM 5.2 cached…

DGX agent

Well said @RamaswmySridhar! I believe the cost saving is actually much bigger, more like 4-5x. E.g., we have just reduced our GLM 5.2 cached token price by 2X to return efficiency gain to our users. A

toolsfireworks-ai--x
26 Jun 2026
Tools

But here's the punchline. Normalized to 90% cache hit rate: GLM-5.2 (Fireworks): 1.12/session Opus-4.7 (Anthropic): 2.14/session GLM is ~4…

DGX agent

This post compares the cost efficiency of Fireworks' GLM-5.2 model versus Anthropic's Opus across cached sessions, showing GLM-5.2 achieving approximately 4x lower cost per session at 90% cache hit ra

toolsfireworks-ai--x
25 Jun 2026
Tools

Congrats! Open source GLM model is really a game changer! Extremely fast, cheap, and high quality!

DGX agent

Fireworks AI announced the release of an open source GLM model that offers significant improvements in speed, cost efficiency, and output quality compared to existing alternatives. The post suggests t

toolsfireworks-ai--x
25 Jun 2026
Model Releases

In a joint Fireworks and @Faros_AI evaluation of 211 real engineering tasks, Claude Code + GLM-5.2 beat both Claude Code + Opus 4.8 and Code…

DGX agent

In a joint Fireworks and @Faros_AI evaluation of 211 real engineering tasks, Claude Code + GLM-5.2 beat both Claude Code + Opus 4.8 and Codex + GPT-5.5: - Judge score: 0.568 vs. 0.521 and 0.466 - Time

model-releasesfireworks-ai--x
25 Jun 2026
Agents

Open + closed models = better together. Our previous research with @harvey showed the benefits of combining a frontier closed model as an ad…

DGX agent

Open + closed models = better together. Our previous research with @harvey showed the benefits of combining a frontier closed model as an advisor agent with fine-tuned, open-source worker agents. Thre

agentsfireworks-ai--x
25 Jun 2026
Model Releases

RL fine-tuning is now live for @nvidiaai Nemotron 3 on Fireworks, starting with Nemotron 3 Super (LoRA). Train with GRPO and serve the model…

DGX agent

RL fine-tuning is now live for @nvidiaai Nemotron 3 on Fireworks, starting with Nemotron 3 Super (LoRA). Train with GRPO and serve the model in one place. We price by GPU-hour, not per token, so long

model-releasesfireworks-ai--x
25 Jun 2026
Model Releases

The underrated part of this announcement is that Fireworks has been quietly great behind the scenes helping us eval and serve these models B…

DGX agent

The underrated part of this announcement is that Fireworks has been quietly great behind the scenes helping us eval and serve these models Big kudos to the Fireworks team! Kimi K2.7 Code and GLM 5.2 a

model-releasesfireworks-ai--x
25 Jun 2026
Applications

GLM 5.2 is the open coding model everyone's been talking about. Now you can fine-tune it on Fireworks. SFT, DPO, and RL all supported. A lea…

DGX agent

GLM 5.2 is the open coding model everyone's been talking about. Now you can fine-tune it on Fireworks. SFT, DPO, and RL all supported. A leaderboard winner can still lose on your codebase. Training cl

applicationsfireworks-ai--x
24 Jun 2026
Tools

It's way easier to switch models than to switch harnesses, and like many of you we use @cursor_ai every day. Now you can try out the latest …

DGX agent

It's way easier to switch models than to switch harnesses, and like many of you we use @cursor_ai every day. Now you can try out the latest open-source frontier model without changing your workflow. Y

toolsfireworks-ai--x
24 Jun 2026
Tools

RT @Prof_OZ: This is worth understanding. Keeping up with AI is hard, I know. It move fast and knowledge compounds. Even I use AI to build…

DGX agent

Dr. Oz discusses the rapid pace of AI development and the challenge of staying informed about advances in the field, noting that even he uses AI tools to help manage and synthesize the growing body of

toolsfireworks-ai--x
24 Jun 2026
Tools

The hard part of reinforcement learning on a frontier model is the infrastructure that keeps training and inference numerically identical: z…

DGX agent

The hard part of reinforcement learning on a frontier model is the infrastructure that keeps training and inference numerically identical: zero KLD, end to end. We've solved this challenge, and are no

toolsfireworks-ai--x
24 Jun 2026
Model Releases

Bring open models to where you're already working. FireConnect brings Fireworks' top models directly into Claude Code, Pi, OpenCode, and Cod…

DGX agent

Bring open models to where you're already working. FireConnect brings Fireworks' top models directly into Claude Code, Pi, OpenCode, and Codex. Watch our Head of AI Education @Prof_oz show you how: Ho

model-releasesfireworks-ai--x
23 Jun 2026
Model Releases

GLM-5.2 has been the most popular new model on Fireworks this past week. @ArtificialAnlys confirms why: #3 overall on GDPval-AA (1524 Elo), …

DGX agent

GLM-5.2 has been the most popular new model on Fireworks this past week. @ArtificialAnlys confirms why: #3 overall on GDPval-AA (1524 Elo), #1 open weights by 116 points. Interest is showing no signs

model-releasesfireworks-ai--x
22 Jun 2026
Agents

You don't need to be AI Dependent to build State of the Art software. @FactoryAI and Fireworks use GLM to build the most advanced agentic AI…

DGX agent

You don't need to be AI Dependent to build State of the Art software. @FactoryAI and Fireworks use GLM to build the most advanced agentic AI you can train and own today. GLM 5.2 is available in Droid,

agentsfireworks-ai--x
22 Jun 2026
Model Releases

An hour in and first impression is definitely that GLM is really solid (very easy to set up on @FireworksAI_HQ, props to them for that, took…

DGX agent

A user shares positive early impressions of GLM (likely a language model), praising its solid performance and ease of setup on Fireworks AI's platform. The post highlights Fireworks AI's developer exp

model-releasesfireworks-ai--x
21 Jun 2026
Tools

also hearing from other customers that GLM 5.2 approaches GPT 5.5 on their evals + use cases

DGX agent

Fireworks AI reports that GLM 5.2, their language model, demonstrates performance comparable to GPT 5.5 across various evaluation benchmarks and real-world use cases. This claim suggests GLM 5.2 is co

toolsfireworks-ai--x
19 Jun 2026
Tools

Coding agents break when models are 'almost' bug-free. But almost valid JSON is just not the same valid JSON. Fun piece here from @akshay_pa…

DGX agent

Coding agents break when models are 'almost' bug-free. But almost valid JSON is just not the same valid JSON. Fun piece here from @akshay_pachaar shows why SFT can't fix this, and how GRPO trains agai

toolsfireworks-ai--x
10 Jun 2026
Model Releases

Fireworks Training Platform keeps expanding. Leading US open weight model Nemotron 3 Ultra is now ready for post-training: SFT and DPO via L…

DGX agent

Fireworks Training Platform keeps expanding. Leading US open weight model Nemotron 3 Ultra is now ready for post-training: SFT and DPO via LoRA or full-parameter, on the same infrastructure that serve

model-releasesfireworks-ai--x
6 Jun 2026
Agents

Fireworks was named to @Redpoint's InfraRed 100 which recognizes the companies building the foundation for the next wave of AI. We're just g…

DGX agent

Fireworks was named to @Redpoint's InfraRed 100 which recognizes the companies building the foundation for the next wave of AI. We're just getting started. Come build with us: https://fireworks.ai/car

agentsfireworks-ai--x
4 Jun 2026
Agents

Many research labs only consider inference efficiency after the fact. Step 3.7 Flash is a 198B sparse MoE VLM designed by @StepFun_ai for in…

DGX agent

Many research labs only consider inference efficiency after the fact. Step 3.7 Flash is a 198B sparse MoE VLM designed by @StepFun_ai for inference from the start. 196B language backbone with a 1.8B v

agentsfireworks-ai--x
4 Jun 2026
Model Releases

NVIDIA Nemotron 3 Ultra is on Fireworks, day zero. Nemotron Ultra is an open model for frontier reasoning and orchestration in long-running …

DGX agent

NVIDIA Nemotron 3 Ultra is on Fireworks, day zero. Nemotron Ultra is an open model for frontier reasoning and orchestration in long-running autonomous agents. Think use cases like coding agents, deep

model-releasesfireworks-ai--x
4 Jun 2026
← Previous
1234
Next →