AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

Search: “fireworks-ai--x”

GridTimelineEvolution
165 results
CompaniesToolsTechniques

Each lane shows up to 8 recent matching entries, ordered from earlier to later. Tracks load separately to keep the 75,000+ entry wiki fast.

Techniques

TechniqueRLHF / Alignment2 recent entries
18 Apr 2026ICYMI from a few weeks back, we compiled our learnings around how to achieve Training-Inference Parity in MoE Models. The Fundamental Issue:…

Fireworks AI published a compilation of insights on achieving training-inference parity in Mixture of Experts (MoE) models, addressing fundamental challenges in aligning model behavior between trainin

→30 Jul 2026one of the top cybersecurity models, post-trained from open weights by @depthfirstlabs on @FireworksAI_HQ long-horizon RL is as much an infr…

one of the top cybersecurity models, post-trained from open weights by @depthfirstlabs on @FireworksAI_HQ long-horizon RL is as much an infra problem as a research one: 100+ turn rollouts, async/pipel

HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
TechniqueRAG1 recent entries
31 Jul 2026Fine-tune your own embedding model for the price of a coffee. A great reranker can't surface a doc that was never retrieved. A RAG pipeline …

Fine-tune your own embedding model for the price of a coffee. A great reranker can't surface a doc that was never retrieved. A RAG pipeline cannot cite a case it failed to retrieve. See how contrastiv

TechniqueAgents8 recent entries
1 Aug 2026DeepSeek-V4-Flash-0731 is live on Fireworks, day-zero. DeepSeek reports it beats V4 Pro across all 9 agentic evals, incl. 82.7% on Terminal …

DeepSeek-V4-Flash-0731 is live on Fireworks, day-zero. DeepSeek reports it beats V4 Pro across all 9 agentic evals, incl. 82.7% on Terminal Bench. Better cost-per-task than V4 Pro, at the economical p

→3 Aug 2026You probably used @FireworksAI_HQ this week. The cool part is that you just did not know you did. 🎆 Open-source models are free, sure... bu…

You probably used @FireworksAI_HQ this week. The cool part is that you just did not know you did. 🎆 Open-source models are free, sure... but the hard part is tailoring them so they perform best on you

→4 Aug 2026If you missed today's K3 on Fireworks webinar, we've got you covered. We'll be partnering with @arizeai to chat more models on Thursday, Aug…

If you missed today's K3 on Fireworks webinar, we've got you covered. We'll be partnering with @arizeai to chat more models on Thursday, August 6th. Sign up below! Next session, August 6 we'll be co l

→5 Aug 2026DeepSeek V4 Flash 0731 is ready to fine-tune on Fireworks. Run SFT, DPO, and RL training jobs on the Dedicated Training API. SFT and DPO fro…

DeepSeek V4 Flash 0731 is ready to fine-tune on Fireworks. Run SFT, DPO, and RL training jobs on the Dedicated Training API. SFT and DPO from the managed UI. Built for coding agents and high-volume pr

→7 Aug 2026In this episode of @wandb's Gradient Dissent , @l2k and Fireworks CEO @lqiao discuss why she believes the industry is at a turning point... …

In this episode of @wandb's Gradient Dissent , @l2k and Fireworks CEO @lqiao discuss why she believes the industry is at a turning point... ... one that calls for more open intelligence, not less. The

→10 Aug 2026Configure and run Fireworks training jobs from your coding agent. Install the Training Skill in one line, describe your goal, and Claude Cod…

Configure and run Fireworks training jobs from your coding agent. Install the Training Skill in one line, describe your goal, and Claude Code, Codex, or Cursor helps choose the method, validate data,

→11 Aug 2026Looking for a faster specialized model for your Agent Work? @nvidia Nemotron 3.5 Lightning (30B MoE, 3B active params) is now live on Firewo…

Looking for a faster specialized model for your Agent Work? @nvidia Nemotron 3.5 Lightning (30B MoE, 3B active params) is now live on Fireworks. It’s distilled from NVIDIA Nemotron 3 Ultra to be your

→12 Aug 2026Muse Glimmer is live on Fireworks. The new open-weight model from Meta Superintelligence Labs is a 30B dense model built for always-on agent…

Muse Glimmer is live on Fireworks. The new open-weight model from Meta Superintelligence Labs is a 30B dense model built for always-on agents that reason across many sequential tool calls and can reco

TechniqueFine-tuning8 recent entries
31 Jul 2026Fine-tune your own embedding model for the price of a coffee. A great reranker can't surface a doc that was never retrieved. A RAG pipeline …

Fine-tune your own embedding model for the price of a coffee. A great reranker can't surface a doc that was never retrieved. A RAG pipeline cannot cite a case it failed to retrieve. See how contrastiv

→1 Aug 2026In a world where building AI applications is getting easier every day, the biggest moat won’t be the application itself. It will be a deep u…

In a world where building AI applications is getting easier every day, the biggest moat won’t be the application itself. It will be a deep understanding of your users. The winners will be the companie

→3 Aug 2026You probably used @FireworksAI_HQ this week. The cool part is that you just did not know you did. 🎆 Open-source models are free, sure... bu…

You probably used @FireworksAI_HQ this week. The cool part is that you just did not know you did. 🎆 Open-source models are free, sure... but the hard part is tailoring them so they perform best on you

→5 Aug 2026DeepSeek V4 Flash 0731 is ready to fine-tune on Fireworks. Run SFT, DPO, and RL training jobs on the Dedicated Training API. SFT and DPO fro…

DeepSeek V4 Flash 0731 is ready to fine-tune on Fireworks. Run SFT, DPO, and RL training jobs on the Dedicated Training API. SFT and DPO from the managed UI. Built for coding agents and high-volume pr

→7 Aug 2026In this episode of @wandb's Gradient Dissent , @l2k and Fireworks CEO @lqiao discuss why she believes the industry is at a turning point... …

In this episode of @wandb's Gradient Dissent , @l2k and Fireworks CEO @lqiao discuss why she believes the industry is at a turning point... ... one that calls for more open intelligence, not less. The

→10 Aug 2026Configure and run Fireworks training jobs from your coding agent. Install the Training Skill in one line, describe your goal, and Claude Cod…

Configure and run Fireworks training jobs from your coding agent. Install the Training Skill in one line, describe your goal, and Claude Code, Codex, or Cursor helps choose the method, validate data,

→11 Aug 2026Leveraging the frontier ecosystem We work with neolabs like @trajectorylabs and @EngramLab and inference + post-training infra providers lik…

Leveraging the frontier ecosystem We work with neolabs like @trajectorylabs and @EngramLab and inference + post-training infra providers like @baseten, @FireworksAI_HQ, and @appliedcompute to post-tra

→12 Aug 2026Muse Glimmer is live on Fireworks. The new open-weight model from Meta Superintelligence Labs is a 30B dense model built for always-on agent…

Muse Glimmer is live on Fireworks. The new open-weight model from Meta Superintelligence Labs is a 30B dense model built for always-on agents that reason across many sequential tool calls and can reco

TechniqueMultimodal5 recent entries
28 Apr 2026Excited to support @NVIDIA Nemotron 3 Nano Omni, now available on Fireworks. It's the first open model that handles vision, audio, video, an…

Excited to support @NVIDIA Nemotron 3 Nano Omni, now available on Fireworks. It's the first open model that handles vision, audio, video, and text in a single inference loop. Built for multimodal sub-

→1 Jun 2026Many research labs only consider inference efficiency after the fact. Step 3.7 Flash is a 196B MoE model, and built for inference from the s…

Many research labs only consider inference efficiency after the fact. Step 3.7 Flash is a 196B MoE model, and built for inference from the start by @StepFun_ai. Multi-Matrix Factorization Attention (M

→4 Jun 2026Many research labs only consider inference efficiency after the fact. Step 3.7 Flash is a 198B sparse MoE VLM designed by @StepFun_ai for in…

Many research labs only consider inference efficiency after the fact. Step 3.7 Flash is a 198B sparse MoE VLM designed by @StepFun_ai for inference from the start. 196B language backbone with a 1.8B v

→15 Jul 2026The best base model for training isn't about someone else's benchmarks. Excited to offer @thinkymachines' first open-weights model: Inkling.…

The best base model for training isn't about someone else's benchmarks. Excited to offer @thinkymachines' first open-weights model: Inkling. 975B MoE, multimodal, with Apache 2.0. A great foundation f

→27 Jul 2026Start owning your own intelligence today with Kimi K3 on Fireworks. Serve it, fine-tune it, and put it into production. What are you waiting…

**Summary:** Kimi K3 is a publicly available 3‑trillion‑parameter language model now live on Fireworks, offering 1 million-token context windows with native vision and reasoning that rivals top closed

TechniqueSafety5 recent entries
28 Apr 2026Prevent prompt injection. safe_tokenization: true Keep your system yours. https://fireworks.ai/blog/safe-tokenization-preventing-prompt-inje…

Safe tokenization is a security feature that helps prevent prompt injection attacks by ensuring that user inputs are properly processed and isolated from system instructions. Fireworks AI discusses ho

→15 Jul 2026Generic AI models are not built for medicine. Doximity Ask, trained and served on Fireworks, outperformed GPT-5.6 Sol, Claude Fable 5, and O…

Generic AI models are not built for medicine. Doximity Ask, trained and served on Fireworks, outperformed GPT-5.6 Sol, Claude Fable 5, and OpenEvidence in a Stanford-Harvard clinical AI safety study.

→24 Jul 2026Open weights = freedom. You can run them on your own hardware. No vendor can pull the plug. No API can deprecate you. No company logs your p…

Open weights = freedom. You can run them on your own hardware. No vendor can pull the plug. No API can deprecate you. No company logs your private data. That's sovereignty. Closed models hand one comp

→27 Jul 2026Nexus connects to the agentic harnesses your teams already use, whether that’s Claude Code, Codex, OpenCode, or your own custom tooling. It …

Nexus connects to the agentic harnesses your teams already use, whether that’s Claude Code, Codex, OpenCode, or your own custom tooling. It gives you: → Intelligent routing, automatically matching eac

→30 Jul 2026one of the top cybersecurity models, post-trained from open weights by @depthfirstlabs on @FireworksAI_HQ long-horizon RL is as much an infr…

one of the top cybersecurity models, post-trained from open weights by @depthfirstlabs on @FireworksAI_HQ long-horizon RL is as much an infra problem as a research one: 100+ turn rollouts, async/pipel