AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
Human
83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

Search: “fireworks-ai--x”

GridTimelineEvolution
61+ results
12 Aug 2026

Muse Glimmer is live on Fireworks. The new open-weight model from Meta Superintelligence Labs is a 30B dense model built for always-on agent…

AgentsDGX agent

Muse Glimmer is live on Fireworks. The new open-weight model from Meta Superintelligence Labs is a 30B dense model built for always-on agents that reason across many sequential tool calls and can reco

11 Aug 2026

Leveraging the frontier ecosystem We work with neolabs like @trajectorylabs and @EngramLab and inference + post-training infra providers lik…

ApplicationsDGX agent

Leveraging the frontier ecosystem We work with neolabs like @trajectorylabs and @EngramLab and inference + post-training infra providers like @baseten, @FireworksAI_HQ, and @appliedcompute to post-tra

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Looking for a faster specialized model for your Agent Work? @nvidia Nemotron 3.5 Lightning (30B MoE, 3B active params) is now live on Firewo…

Model ReleasesDGX agent

Looking for a faster specialized model for your Agent Work? @nvidia Nemotron 3.5 Lightning (30B MoE, 3B active params) is now live on Fireworks. It’s distilled from NVIDIA Nemotron 3 Ultra to be your

10 Aug 2026

Configure and run Fireworks training jobs from your coding agent. Install the Training Skill in one line, describe your goal, and Claude Cod…

Model ReleasesDGX agent

Configure and run Fireworks training jobs from your coding agent. Install the Training Skill in one line, describe your goal, and Claude Code, Codex, or Cursor helps choose the method, validate data,

Voyage AI models now run natively on Fireworks, the first and only dedicated inference platform @VoyageAI by @MongoDB has partnered with. Em…

ToolsDGX agent

Voyage AI models now run natively on Fireworks, the first and only dedicated inference platform @VoyageAI by @MongoDB has partnered with. Embed, retrieve, rerank, generate: your full retrieval pipelin

7 Aug 2026

In this episode of @wandb's Gradient Dissent , @l2k and Fireworks CEO @lqiao discuss why she believes the industry is at a turning point... …

Model ReleasesDGX agent

In this episode of @wandb's Gradient Dissent , @l2k and Fireworks CEO @lqiao discuss why she believes the industry is at a turning point... ... one that calls for more open intelligence, not less. The

6 Aug 2026

This webinar is happening in 30 minutes, and that means there's still time to register! Following the discussion will be an open Q&A with ou…

Model ReleasesDGX agent

This webinar is happening in 30 minutes, and that means there's still time to register! Following the discussion will be an open Q&A with our Head of AI Education @Prof_OZ, and the @arizeai team. See

5 Aug 2026

DeepSeek V4 Flash 0731 is ready to fine-tune on Fireworks. Run SFT, DPO, and RL training jobs on the Dedicated Training API. SFT and DPO fro…

Model ReleasesDGX agent

DeepSeek V4 Flash 0731 is ready to fine-tune on Fireworks. Run SFT, DPO, and RL training jobs on the Dedicated Training API. SFT and DPO from the managed UI. Built for coding agents and high-volume pr

4 Aug 2026

If you missed today's K3 on Fireworks webinar, we've got you covered. We'll be partnering with @arizeai to chat more models on Thursday, Aug…

Model ReleasesDGX agent

If you missed today's K3 on Fireworks webinar, we've got you covered. We'll be partnering with @arizeai to chat more models on Thursday, August 6th. Sign up below! Next session, August 6 we'll be co l

Remember OpenAI tripled its score on ARC-AGI with a harness fix? We fixed harness for Kimi K3 on CyberGym E2E: 2x better at vulnerability de…

Model ReleasesDGX agent

Remember OpenAI tripled its score on ARC-AGI with a harness fix? We fixed harness for Kimi K3 on CyberGym E2E: 2x better at vulnerability detection, 3x at patching K3 is SOTA cyber-defense model you c

Running fast is not enough, you need fast AND correct An excellent addition from @ArtificialAnlys to make sure that the flashy speed numbers…

Model ReleasesDGX agent

Running fast is not enough, you need fast AND correct An excellent addition from @ArtificialAnlys to make sure that the flashy speed numbers are backed by 100% matching accuracy Announcing the Artific

3 Aug 2026

Most companies still rent AI by the token, build on someone else’s roadmap, and hope the next model release does not disrupt their systems. …

ApplicationsDGX agent

Most companies still rent AI by the token, build on someone else’s roadmap, and hope the next model release does not disrupt their systems. At ODSC AI West 2026, @Prof_OZ, Head of AI Developer Educati

You probably used @FireworksAI_HQ this week. The cool part is that you just did not know you did. 🎆 Open-source models are free, sure... bu…

Model ReleasesDGX agent

You probably used @FireworksAI_HQ this week. The cool part is that you just did not know you did. 🎆 Open-source models are free, sure... but the hard part is tailoring them so they perform best on you

2 Aug 2026

Cybersecurity isn’t a fortress problem, it’s an immunity problem. Think vaccines. Eliminating pathogen is not practically possible. Vaccines…

Model ReleasesDGX agent

Cybersecurity isn’t a fortress problem, it’s an immunity problem. Think vaccines. Eliminating pathogen is not practically possible. Vaccines don’t eliminate pathogens. They teach the immune system to

1 Aug 2026

DeepSeek-V4-Flash-0731 is live on Fireworks, day-zero. DeepSeek reports it beats V4 Pro across all 9 agentic evals, incl. 82.7% on Terminal …

Model ReleasesDGX agent

DeepSeek-V4-Flash-0731 is live on Fireworks, day-zero. DeepSeek reports it beats V4 Pro across all 9 agentic evals, incl. 82.7% on Terminal Bench. Better cost-per-task than V4 Pro, at the economical p

In a world where building AI applications is getting easier every day, the biggest moat won’t be the application itself. It will be a deep u…

ToolsDGX agent

In a world where building AI applications is getting easier every day, the biggest moat won’t be the application itself. It will be a deep understanding of your users. The winners will be the companie

31 Jul 2026

Fine-tune your own embedding model for the price of a coffee. A great reranker can't surface a doc that was never retrieved. A RAG pipeline …

ToolsDGX agent

Fine-tune your own embedding model for the price of a coffee. A great reranker can't surface a doc that was never retrieved. A RAG pipeline cannot cite a case it failed to retrieve. See how contrastiv

If LoRA is underperforming, don't reach for more expensive full parameter fine-tuning right away. We ran three cheap tests (data coverage, o…

Model ReleasesDGX agent

If LoRA is underperforming, don't reach for more expensive full parameter fine-tuning right away. We ran three cheap tests (data coverage, optimization, rank) to see if we could close the gap between

Period.

Model ReleasesDGX agent

Period. If LoRA is underperforming, don't reach for more expensive full parameter fine-tuning right away. We ran three cheap tests (data coverage, optimization, rank) to see if we could close the gap

Security at the foundation requires openness at the foundation. As open-weight models become critical infrastructure for the next generation…

HardwareDGX agent

Security at the foundation requires openness at the foundation. As open-weight models become critical infrastructure for the next generation of software, the systems around them must be transparent, i

30 Jul 2026

Excited to be on the CNBC live show!

ToolsDGX agent

Excited to be on the CNBC live show! Back from vacation and LIVE at 12pm PT / 3pm ET Is AI’s easy-money era ending? We’ll unpack a wild week for the AI trade—big tech earnings, Leopold Aschenbrenner’s

I’m excited to share my next chapter: I’ve joined @FireworksAI_HQ . From AMD, Apple, Uber, Meta, Google, and most recently Snowflake, I’ve w…

ApplicationsDGX agent

I’m excited to share my next chapter: I’ve joined @FireworksAI_HQ . From AMD, Apple, Uber, Meta, Google, and most recently Snowflake, I’ve worked on many of the foundational technologies that power mo

one of the top cybersecurity models, post-trained from open weights by @depthfirstlabs on @FireworksAI_HQ long-horizon RL is as much an infr…

SafetyDGX agent

one of the top cybersecurity models, post-trained from open weights by @depthfirstlabs on @FireworksAI_HQ long-horizon RL is as much an infra problem as a research one: 100+ turn rollouts, async/pipel

Our partners at @depthfirstlabs just released dfs-large1, a specialized model built for finding and validating real vulnerabilities in large…

Model ReleasesDGX agent

Our partners at @depthfirstlabs just released dfs-large1, a specialized model built for finding and validating real vulnerabilities in large enterprise codebases. We helped them to scale the training

29 Jul 2026

DeepSeek V4 Flash isn't just for inference anymore. Fine-tune it on Fireworks with supervised fine-tuning, preference tuning, and combined p…

Model ReleasesDGX agent

DeepSeek V4 Flash isn't just for inference anymore. Fine-tune it on Fireworks with supervised fine-tuning, preference tuning, and combined preference optimization from the managed UI. Reinforcement le

28 Jul 2026

Many AI tools rent intelligence. Kimi K3 on Fireworks is different. Fine-tune it on your own data, serve from US-hosted endpoints, and own t…

ToolsDGX agent

Many AI tools rent intelligence. Kimi K3 on Fireworks is different. Fine-tune it on your own data, serve from US-hosted endpoints, and own the weights. Zero data retention. ICYMI yesterday, start buil

Our users have been really excited to try Kimi K3 and we're excited to bring it to Augment's product family. It is the most capable open-sou…

AgentsDGX agent

Our users have been really excited to try Kimi K3 and we're excited to bring it to Augment's product family. It is the most capable open-source model that we have tested to date. Try it today in Cosmo

You can now fine-tune Kimi K3 on Fireworks. Conduct supervised fine-tuning, preference tuning, and reinforcement learning via Training API. …

ToolsDGX agent

You can now fine-tune Kimi K3 on Fireworks. Conduct supervised fine-tuning, preference tuning, and reinforcement learning via Training API. Run across dedicated, and serverless training. The first ope

27 Jul 2026

fyi @FireworksAI_HQ launched inference AND training for K3 on day 0. having 'specialized intelligence' isn't just inference, it's inference …

ToolsDGX agent

fyi @FireworksAI_HQ launched inference AND training for K3 on day 0. having 'specialized intelligence' isn't just inference, it's inference + training, together. That's why K3 on Fireworks offers both

Kimi K3 is live on Fireworks. Day 0, inference and training. US-hosted, and zero data retention. This is the first frontier open model in th…

Model ReleasesDGX agent

Kimi K3 is live on Fireworks. Day 0, inference and training. US-hosted, and zero data retention. This is the first frontier open model in the 3 trillion parameter class. It sports 1M context, native v

Kimi K3 is live on @FireworksAI_HQ As per usual not only are we fast AF but we take quality as a P0 and never sacrifice quality for anything…

ApplicationsDGX agent

Kimi K3 is live on @FireworksAI_HQ As per usual not only are we fast AF but we take quality as a P0 and never sacrifice quality for anything (only OSS provider to top Moonshot's evals across the board

Kimi K3 rivals Anthropic and OpenAI’s top models at a fraction of the cost. 3 / 1M input tokens. 15 / 1M output. $0.30 / 1M cached You can…

ToolsDGX agent

Kimi K3 rivals Anthropic and OpenAI’s top models at a fraction of the cost. 3 / 1M input tokens. 15 / 1M output. $0.30 / 1M cached You can now route your hardest reasoning to an open model without pay

Make Kimi K3 yours with LoRA training on Fireworks Training a model this massive used to be a big project. With Fireworks Training, you can …

HardwareDGX agent

Make Kimi K3 yours with LoRA training on Fireworks Training a model this massive used to be a big project. With Fireworks Training, you can take the best open model in the world and efficiently tune i

Nexus connects to the agentic harnesses your teams already use, whether that’s Claude Code, Codex, OpenCode, or your own custom tooling. It …

Model ReleasesDGX agent

Nexus connects to the agentic harnesses your teams already use, whether that’s Claude Code, Codex, OpenCode, or your own custom tooling. It gives you: → Intelligent routing, automatically matching eac

Start owning your own intelligence today with Kimi K3 on Fireworks. Serve it, fine-tune it, and put it into production. What are you waiting…

ApplicationsDGX agent

**Summary:** Kimi K3 is a publicly available 3‑trillion‑parameter language model now live on Fireworks, offering 1 million-token context windows with native vision and reasoning that rivals top closed

24 Jul 2026

Open weights = freedom. You can run them on your own hardware. No vendor can pull the plug. No API can deprecate you. No company logs your p…

Local AiDGX agent

Open weights = freedom. You can run them on your own hardware. No vendor can pull the plug. No API can deprecate you. No company logs your private data. That's sovereignty. Closed models hand one comp

23 Jul 2026

@HarryStebbings @lqiao Spotify https://open.spotify.com/episode/14rh372tSdEzRITBQz9HSP?si=c9ed4a9eae9f4de0 Youtube https://youtu.be/PCAiqKCf…

ToolsDGX agent

@HarryStebbings @lqiao Spotify https://open.spotify.com/episode/14rh372tSdEzRITBQz9HSP?si=c9ed4a9eae9f4de0 Youtube https://youtu.be/PCAiqKCfRSk?si=WDP2PfIkdn0XYH9T Apple Podcasts https://podcasts.appl

Many still debate open vs closed, and compare cost per token (accounting metric). Better to shift attention to cost per successful task. @se…

Model ReleasesDGX agent

Many still debate open vs closed, and compare cost per token (accounting metric). Better to shift attention to cost per successful task. @seldo and the team at @arizeai did so across 2,400 runs. Concl

21 Jul 2026

The MiniMax M3 model is now available for training on Fireworks! You can use managed LoRA SFT and DPO for standard fine-tuning workflows, or…

ToolsDGX agent

The MiniMax M3 model is now available for training on Fireworks! You can use managed LoRA SFT and DPO for standard fine-tuning workflows, or the Fireworks Training API for custom SFT, DPO, and RL loop

The team @tryheidi didn't want to keep renting someone else's intelligence. So Heidi fine-tuned an open model that beat Gemini Pro on qualit…

Model ReleasesDGX agent

The team @tryheidi didn't want to keep renting someone else's intelligence. So Heidi fine-tuned an open model that beat Gemini Pro on quality in their internal evals, and ran with 3.5x faster latency

We'll see you in an hour!

ToolsDGX agent

We'll see you in an hour! Want to chat live with me about fine-tuning, Kimi K3, or really anything else AI? Come to our first of many @FireworksAI_HQ office hours tomorrow @ 10am PT See you there! htt

20 Jul 2026

Our first DevRel office hours kicks off tomorrow. Come chat with our Head of AI Education @Prof_OZ about anything and everything AI.

ApplicationsDGX agent

Our first DevRel office hours kicks off tomorrow. Come chat with our Head of AI Education @Prof_OZ about anything and everything AI. Want to chat live with me about fine-tuning, Kimi K3, or really any

15 Jul 2026

Generic AI models are not built for medicine. Doximity Ask, trained and served on Fireworks, outperformed GPT-5.6 Sol, Claude Fable 5, and O…

Model ReleasesDGX agent

Generic AI models are not built for medicine. Doximity Ask, trained and served on Fireworks, outperformed GPT-5.6 Sol, Claude Fable 5, and OpenEvidence in a Stanford-Harvard clinical AI safety study.

The best base model for training isn't about someone else's benchmarks. Excited to offer @thinkymachines' first open-weights model: Inkling.…

ToolsDGX agent

The best base model for training isn't about someone else's benchmarks. Excited to offer @thinkymachines' first open-weights model: Inkling. 975B MoE, multimodal, with Apache 2.0. A great foundation f

Trending AND fastest-growing in the same month? Benchmarks change daily, but only @tryramp has the database of real receipts to track this. …

Model ReleasesDGX agent

Trending AND fastest-growing in the same month? Benchmarks change daily, but only @tryramp has the database of real receipts to track this. We can confirm: demand for open-weight inference and trainin

14 Jul 2026

Last month we hosted an AI Nerd Meetup @Workato HQ featuring builders tackling the hardest problems in agentic AI. Highlights: @yuan_sihan o…

AgentsDGX agent

Last month we hosted an AI Nerd Meetup @Workato HQ featuring builders tackling the hardest problems in agentic AI. Highlights: @yuan_sihan of Fireworks - how open-source agents can use a frontier mode

Our free monthly DevRel webinar series kicks off this Thursday, 10am PST. Fine-tune an open vision model to extract clean, structured JSON f…

ToolsDGX agent

Our free monthly DevRel webinar series kicks off this Thursday, 10am PST. Fine-tune an open vision model to extract clean, structured JSON from messy receipt images, managed start to finish on Firewor

10 Jul 2026

Long-context sparse attention has a catch: data-dependent block selection wrecks memory access kills speed. Our @MiniMax_AI M3 kernel on Bla…

HardwareDGX agent

Long-context sparse attention has a catch: data-dependent block selection wrecks memory access kills speed. Our @MiniMax_AI M3 kernel on Blackwell answers it. KV-stationary, each block read once, ~980

9 Jul 2026

Loving this piece: bringing delight to your daily coffee, your daily workflows, and your AI bill. @gumloop is now running open-weight models…

ApplicationsDGX agent

Loving this piece: bringing delight to your daily coffee, your daily workflows, and your AI bill. @gumloop is now running open-weight models on Fireworks in production. 80% lower cost vs. closed-lab A

8 Jul 2026

Happy to announce that we’ve enlisted @FireworksAI_HQ as our primary provider for open-source models. In AI inference, quality, latency, rel…

ToolsDGX agent

Happy to announce that we’ve enlisted @FireworksAI_HQ as our primary provider for open-source models. In AI inference, quality, latency, reliability, and privacy vary a lot. Fireworks has earned our t

If you're tired of re-explaining the same context to your AI every day, go see what @PromptQL just launched!

ToolsDGX agent

If you're tired of re-explaining the same context to your AI every day, go see what @PromptQL just launched! We raised $136M to kill Slack. Introducing PromptQL: The first AI version of Slack. Here’s

Not another demo or benchmark. @ShoucongChen is a senior member of our technical staff. A real project, scoped at 1-month. Delivered in 4 da…

Model ReleasesDGX agent

Not another demo or benchmark. @ShoucongChen is a senior member of our technical staff. A real project, scoped at 1-month. Delivered in 4 days with GLM5.2 Fast. The best devs deserve >400 t/sec. Take

This is what production RL infra looks like. ICYMI: RL training at scale separates into two distinct problems. - Tight collective comms for …

ApplicationsDGX agent

This is what production RL infra looks like. ICYMI: RL training at scale separates into two distinct problems. - Tight collective comms for the trainer. - Distributed async inference for rollout. Kudo

7 Jul 2026

Attending @RaiseSummit? Join our CEO @lqiao tomorrow @ 4:00 PM for a fireside chat with @MattEvantic of Evantic Capital on the Master Stage.…

ToolsDGX agent

Fireworks AI is hosting a fireside chat at Raise Summit featuring CEO Lqiao and Matt Evantic from Evantic Capital on the Master Stage at 4:00 PM. The event appears to be a discussion panel or intervie

What actually makes an AI application good or not? Thanks to @StackOverflow for hosting @the_bunny_chen on the Stack Overflow Podcast (even …

TutorialsDGX agent

What actually makes an AI application good or not? Thanks to @StackOverflow for hosting @the_bunny_chen on the Stack Overflow Podcast (even though he is giving away all of our secrets!) Listen here: h

6 Jul 2026

Today, we're helping kick off AMD AI Developer Hackathon: ACT II. We're giving $50 in Fireworks credits to participants in this event which …

ToolsDGX agent

Today, we're helping kick off AMD AI Developer Hackathon: ACT II. We're giving 50 in Fireworks credits to participants in this event which also includes: 20K in prizes. $150 in credits. Fully online.

3 Jul 2026

Don't forget, we'll be in Paris for @RaiseSummit. Will we see you there?

HardwareDGX agent

Don't forget, we'll be in Paris for @RaiseSummit. Will we see you there? Going to be in Paris for @RaiseSummit? Join us and @nvidia to hear from leaders at both companies on the future of AI inference

Prompt engineering is costing you money. Learn how to fine-tune your models. Stuffing a lot of text into every single API call slows down yo…

AgentsDGX agent

Prompt engineering is costing you money. Learn how to fine-tune your models. Stuffing a lot of text into every single API call slows down your app because the model has to process all those tokens bef

2 Jul 2026

Fear not, @FireworksAI_HQ is perfectly legal and serves the best models on 🇺🇸 servers for your AI-ndependence

ApplicationsDGX agent

Fear not, @FireworksAI_HQ is perfectly legal and serves the best models on 🇺🇸 servers for your AI-ndependence For info on firework enforcement, park closures and the like for this Fourth of July weeke

1 Jul 2026

do you know what you pay for in agentic workloads? cached tokens! session with 50+ tool calls -> prompt is billed 50 times all providers giv…

Model ReleasesDGX agent

do you know what you pay for in agentic workloads? cached tokens! session with 50+ tool calls -> prompt is billed 50 times all providers give 1/5 cached discount for GLM-5.2 we at @FireworksAI_HQ drop

← Previous
123
Next →
165 results