AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

Search: “fireworks-ai--x”

GridTimelineEvolution
165 results
Tools

We spent the week at #MSBuild talking about one thing: fine-tuning has gone from 'maybe not worth it' to your actual competitive moat. @lqia…

DGX agent

We spent the week at #MSBuild talking about one thing: fine-tuning has gone from 'maybe not worth it' to your actual competitive moat. @lqiao sat down with @yina_arenas to break down why and what Fire

toolsfireworks-ai--x
4 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Applications

Day 2 at #MSBuild is about what it takes to move beyond generic foundation models. Think customization, inference performance, and getting p…

DGX agent

Day 2 at #MSBuild is about what it takes to move beyond generic foundation models. Think customization, inference performance, and getting production-ready AI deployed at scale. @chahvivi will lead a

applicationsfireworks-ai--x
3 Jun 2026
Applications

Fine-tuning to production inference is the gap where teams get stuck. At #MSBuild today, our own Rob Ferguson, @danielhanchen (@UnslothAI) a…

DGX agent

Fine-tuning to production inference is the gap where teams get stuck. At #MSBuild today, our own Rob Ferguson, @danielhanchen (@UnslothAI) and @marksaroufim (@coreautoai) discuss: model customization

applicationsfireworks-ai--x
3 Jun 2026
Model Releases

Frontier models are powerful advisors. On @harvey's Legal Agent Benchmark, a GLM 5.1 worker using Claude Opus 4.7 as a sparse advisor reache…

DGX agent

Frontier models are powerful advisors. On @harvey's Legal Agent Benchmark, a GLM 5.1 worker using Claude Opus 4.7 as a sparse advisor reached 18/100 all-pass versus 14/100 for Opus alone, at 39% of th

model-releasesfireworks-ai--x
3 Jun 2026
Model Releases

MiniMax M3 arrives with MiniMax Sparse Attention (MSA), 15.6x faster decoding at 1M tokens. We're partnering with @MiniMax_AI to power the i…

DGX agent

MiniMax M3 arrives with MiniMax Sparse Attention (MSA), 15.6x faster decoding at 1M tokens. We're partnering with @MiniMax_AI to power the inference behind this week's launch. Head to http://minimax.i

model-releasesfireworks-ai--x
3 Jun 2026
Tutorials

Working with @FireworksAI_HQ to make MAI models easy to fine-tune and fully yours.

DGX agent

Working with @FireworksAI_HQ to make MAI models easy to fine-tune and fully yours. Microsoft MAI models. Coming soon to Fireworks. Intelligence you control. End-to-end lineage you can prove. Fine-tune

tutorialsfireworks-ai--x
3 Jun 2026
Tutorials

Microsoft MAI models. Coming soon to Fireworks. Intelligence you control. End-to-end lineage you can prove. Fine-tune MAI reasoning models f…

DGX agent

Microsoft MAI models. Coming soon to Fireworks. Intelligence you control. End-to-end lineage you can prove. Fine-tune MAI reasoning models for your enterprise tasks. Your data. Your custom models. You

tutorialsfireworks-ai--x
2 Jun 2026
Applications

Move from test to production by running high-performance inference directly on Foundry. At #MSBuild, we demoed an end-to-end workflow showin…

DGX agent

Move from test to production by running high-performance inference directly on Foundry. At #MSBuild, we demoed an end-to-end workflow showing how unified infrastructure improves latency, reduces cost,

applicationsfireworks-ai--x
2 Jun 2026
Model Releases

Super excited to announce seven new world-class MAI models today. They represent what we consider a new era in AI designed to keep you in co…

DGX agent

Super excited to announce seven new world-class MAI models today. They represent what we consider a new era in AI designed to keep you in control and on the frontier. First is our text foundation mode

model-releasesfireworks-ai--x
2 Jun 2026
Tools

We’re looking forward to seeing how developers and enterprises use Fireworks AI on @Microsoft Foundry to power the next generation of intell…

DGX agent

We’re looking forward to seeing how developers and enterprises use Fireworks AI on @Microsoft Foundry to power the next generation of intelligent applications. Catch us at booth F111 at #MSBuild and s

toolsfireworks-ai--x
2 Jun 2026
Model Releases

Many research labs only consider inference efficiency after the fact. Step 3.7 Flash is a 196B MoE model, and built for inference from the s…

DGX agent

Many research labs only consider inference efficiency after the fact. Step 3.7 Flash is a 196B MoE model, and built for inference from the start by @StepFun_ai. Multi-Matrix Factorization Attention (M

model-releasesfireworks-ai--x
1 Jun 2026
Applications

Production AI systems place very different demands on infrastructure once workloads scale. How? Join us at #MSBuild to find out. Register he…

DGX agent

Production AI systems require fundamentally different infrastructure approaches as workloads scale, with demands that differ significantly from development or testing environments. Fireworks AI discus

applicationsfireworks-ai--x
1 Jun 2026
Model Releases

Another proof point for the open-weights thesis. From @RampLabs: 'If we built this again, we'd lean more on open-weight models.' Ramp pointe…

DGX agent

Another proof point for the open-weights thesis. From @RampLabs: 'If we built this again, we'd lean more on open-weight models.' Ramp pointed 10K agents at their own backend. Kimi K2.6 and DeepSeek V4

model-releasesfireworks-ai--x
29 May 2026
Tools

Reliability shouldn't require reserving GPUs. Serverless 2.0 is live on Fireworks: one API, 3 serving paths. → Standard: elastic default → P…

DGX agent

Reliability shouldn't require reserving GPUs. Serverless 2.0 is live on Fireworks: one API, 3 serving paths. → Standard: elastic default → Priority: sheds last under congestion, pricing ~1.5x standard

toolsfireworks-ai--x
29 May 2026
Tools

This tracks. 30 trillion tokens a day on our end, and open model share keeps climbing. Our partners @FactoryAI are seeing what we're seeing.

DGX agent

This tracks. 30 trillion tokens a day on our end, and open model share keeps climbing. Our partners @FactoryAI are seeing what we're seeing. Narrative violation: Open model use in Factory has more tha

toolsfireworks-ai--x
28 May 2026
Tools

10/ The bigger point: your product is the best RL environment you'll ever have. Frontier labs ship models that are good at everything. The o…

DGX agent

10/ The bigger point: your product is the best RL environment you'll ever have. Frontier labs ship models that are good at everything. The opportunity is a model that's great at your thing. Product, u

toolsfireworks-ai--x
27 May 2026
Tools

9/ Real-time RL is where it gets fun. Catch live signals from real users on real generations. Update continuously. Ship a new version every …

DGX agent

9/ Real-time RL is where it gets fun. Catch live signals from real users on real generations. Update continuously. Ship a new version every few hours. Only works if the base model is already good enou

toolsfireworks-ai--x
27 May 2026
Tools

Fireworks is coming to Tech Week, and we're doing it across multiple cities. Boston is up first. We're co-hosting a Founder Skybar Social wi…

DGX agent

Fireworks is coming to Tech Week, and we're doing it across multiple cities. Boston is up first. We're co-hosting a Founder Skybar Social with @fin_ai and @TrustVanta. Follow along via #BOSTechWeek In

toolsfireworks-ai--x
26 May 2026
Model Releases

Cursor's new Composer 2.5 takes third on the Artificial Analysis Coding Agent Index and is ~10-60x lower cost than the higher-effort Opus 4.…

DGX agent

Cursor's new Composer 2.5 takes third on the Artificial Analysis Coding Agent Index and is ~10-60x lower cost than the higher-effort Opus 4.7 and GPT-5.5 variants above it. This release puts Composer

model-releasesfireworks-ai--x
21 May 2026
Hardware

Fine-tuning used to mean a team, a GPU cluster, and weeks of iteration. Now it's just a CLI command, ~10 min of GPU time, a few cents of com…

DGX agent

Fine-tuning used to mean a team, a GPU cluster, and weeks of iteration. Now it's just a CLI command, ~10 min of GPU time, a few cents of compute. You walk away owning the weights. Open models off the

hardwarefireworks-ai--x
21 May 2026
Tutorials

Nathan's @cursor_ai team didn't prompt-engineer their way to Composer 2.5. They trained it. The massive RL program runs RL rollouts on Firew…

DGX agent

Nathan's @cursor_ai team didn't prompt-engineer their way to Composer 2.5. They trained it. The massive RL program runs RL rollouts on Fireworks, alongside production inference. 'Comment 🔥 to see my p

tutorialsfireworks-ai--x
21 May 2026
Tools

Fireworks is coming to Tech Week First up: Boston. We're co-hosting a rooftop founder social with @fin_ai and @TrustVanta — curated crowd, r…

DGX agent

Fireworks is coming to Tech Week First up: Boston. We're co-hosting a rooftop founder social with @fin_ai and @TrustVanta — curated crowd, real conversations, limited space. Thu May 28 · 7–10pm · #BOS

toolsfireworks-ai--x
20 May 2026
Agents

We ran 720 browser agent tasks with @nottecore across frontier models. One baseline model produced malformed outputs in ~1 out of every 5 ca…

DGX agent

We ran 720 browser agent tasks with @nottecore across frontier models. One baseline model produced malformed outputs in ~1 out of every 5 calls, leading to retries inside multi-step workflows. Across

agentsfireworks-ai--x
20 May 2026
Tutorials

Not all the good stuff at Build happens on the main stage. Join Microsoft for Startups on June 2 for Dev Your Own Way. Hands-on activations …

DGX agent

Not all the good stuff at Build happens on the main stage. Join Microsoft for Startups on June 2 for Dev Your Own Way. Hands-on activations with @GitHub, @FireworksAI_HQ and @NVIDIAforStartups. Real c

tutorialsfireworks-ai--x
18 May 2026
Tutorials

The @cursor_ai team shipped Composer 2 and now Composer 2.5 on the same Kimi K2.5 base model. Performance benchmarks are📈. Frontier quality…

DGX agent

The @cursor_ai team shipped Composer 2 and now Composer 2.5 on the same Kimi K2.5 base model. Performance benchmarks are📈. Frontier quality and open-source economics. 85% of the compute powering these

tutorialsfireworks-ai--x
18 May 2026
Model Releases

a new era of hackathons: 2005→ look what i built with web search 2016 → can you build a rails app in 24 hrs? 2025 → spin up a CX bot with Cl…

DGX agent

a new era of hackathons: 2005→ look what i built with web search 2016 → can you build a rails app in 24 hrs? 2025 → spin up a CX bot with Claude in 5 min 2026 → train your own model over a weekend hap

model-releasesfireworks-ai--x
17 May 2026
Tools

Cool billboard @FireworksAI_HQ

DGX agent

Fireworks AI posted about a billboard at their headquarters, likely showcasing their company branding, a product announcement, or marketing campaign. The post appears to be promotional content shared

toolsfireworks-ai--x
15 May 2026
Tools

free to start. fast on day zero. proud to power the default model for LangSmith Fleet. happy building.

DGX agent

free to start. fast on day zero. proud to power the default model for LangSmith Fleet. happy building. LangSmith Fleet now has a free model powered by @FireworksAI_HQ for Developer and Plus plans. It’

toolsfireworks-ai--x
15 May 2026
Model Releases

Weekends are for vibe coding. But are your vibes continuously improving? Fine-tune your own model → stop waiting on someone else's release c…

DGX agent

Weekends are for vibe coding. But are your vibes continuously improving? Fine-tune your own model → stop waiting on someone else's release cycle. Today's training update: Gemma 4 Dense is now availabl

model-releasesfireworks-ai--x
15 May 2026
Model Releases

Fireworks Training Platform continues to expand. Today GLM 5.1 LoRA RL is now live via Training API: SFT, DPO, and full RL on a 200K context…

DGX agent

Fireworks Training Platform continues to expand. Today GLM 5.1 LoRA RL is now live via Training API: SFT, DPO, and full RL on a 200K context window → custom loss functions or smart defaults. No usage

model-releasesfireworks-ai--x
14 May 2026
Tutorials

If you're calling a third-party API, your competitor can make the same call tomorrow. @lqiao is the keynote speaker at @pycon tomorrow, brea…

DGX agent

If you're calling a third-party API, your competitor can make the same call tomorrow. @lqiao is the keynote speaker at @pycon tomorrow, breaking down how to build a durable moat using fine-tuned model

tutorialsfireworks-ai--x
14 May 2026
Model Releases

Kimi K2.6 and DeepSeek V4 Pro are now GA on @FireworksAI_HQ on Foundry + PTU support in the US Data Zone—predictable performance, data resid…

DGX agent

Kimi K2.6 and DeepSeek V4 Pro are now GA on @FireworksAI_HQ on Foundry + PTU support in the US Data Zone—predictable performance, data residency, enterprise SLAs. Frontier open models. Azure controls.

model-releasesfireworks-ai--x
14 May 2026
Model Releases

Like any AI dev not employed by a closed lab, we share the ambition that at least 10x more should be capable of training frontier models in …

DGX agent

Like any AI dev not employed by a closed lab, we share the ambition that at least 10x more should be capable of training frontier models in 2026. And like we all know, some 10x leaps can't survive a c

model-releasesfireworks-ai--x
14 May 2026
Tutorials

Most teams can pick frontier models. Fewer can run them at production scale without hitting constraints in latency, throughput, and governan…

DGX agent

Most teams can pick frontier models. Fewer can run them at production scale without hitting constraints in latency, throughput, and governance. Fireworks AI on @Azure AI Foundry provides the inference

tutorialsfireworks-ai--x
14 May 2026
Tools

𝐅𝐮𝐥𝐥-𝐏𝐚𝐫𝐚𝐦 𝐑𝐋 𝐧𝐨𝐰 𝐚𝐯𝐚𝐢𝐥𝐚𝐛𝐥𝐞 𝐟𝐨𝐫 𝐊𝐢𝐦𝐢 𝐊𝟐.𝟔 You've been told only 3 AI labs matter. The best AI apps never be…

DGX agent

𝐅𝐮𝐥𝐥-𝐏𝐚𝐫𝐚𝐦 𝐑𝐋 𝐧𝐨𝐰 𝐚𝐯𝐚𝐢𝐥𝐚𝐛𝐥𝐞 𝐟𝐨𝐫 𝐊𝐢𝐦𝐢 𝐊𝟐.𝟔 You've been told only 3 AI labs matter. The best AI apps never believed that. @cursor_ai, @vercel, @genspark_ai don't run only off-the-shelf models. They trai

toolsfireworks-ai--x
12 May 2026
Model Releases

Fine-tuning on your proprietary data is the highest leverage thing you can do. Prompts get copied overnight. A model trained on your data, y…

DGX agent

Fine-tuning on your proprietary data is the highest leverage thing you can do. Prompts get copied overnight. A model trained on your data, your evals, your edge cases is a strong moat. OpenAI is windi

model-releasesfireworks-ai--x
11 May 2026
Tools

Frontier labs are betting AGI models will be so good you won't ever want to customize them. We think different. Building on a closed platfor…

DGX agent

Frontier labs are betting AGI models will be so good you won't ever want to customize them. We think different. Building on a closed platform means renting your intelligence. The landlord sets the ter

toolsfireworks-ai--x
9 May 2026
Tutorials

Learn how to structure your AI stack. Our co-founder @the_bunny_chen talks with @theinfrapod's @tnachen and @ianlivingstone about inference …

DGX agent

Learn how to structure your AI stack. Our co-founder @the_bunny_chen talks with @theinfrapod's @tnachen and @ianlivingstone about inference optimization, the open vs. closed model debate, and building

tutorialsfireworks-ai--x
7 May 2026
Model Releases

We’ve been working closely with the @harvey team on the launch of the Legal Agent Benchmark, a product focused on evaluating how open-weight…

DGX agent

We’ve been working closely with the @harvey team on the launch of the Legal Agent Benchmark, a product focused on evaluating how open-weight models perform on long-horizon, real-world legal tasks. Che

model-releasesfireworks-ai--x
6 May 2026
Tools

Why are four of the top 10 trending vendors AI inference platforms? Companies are getting selective about which model handles which task, an…

DGX agent

Why are four of the top 10 trending vendors AI inference platforms? Companies are getting selective about which model handles which task, and routing tokens accordingly. Read more from @arakharazian a

toolsfireworks-ai--x
6 May 2026
Tools

🤝 OpenHuman x Fireworks AI We are excited to announce one of our very first partnerships with Fireworks AI. OpenHuman is now powered by @Fi…

DGX agent

🤝 OpenHuman x Fireworks AI We are excited to announce one of our very first partnerships with Fireworks AI. OpenHuman is now powered by @FireworksAI_HQ: the fastest inference for Qwen3 outside of Alib

toolsfireworks-ai--x
4 May 2026
Applications

Our cofounder @the_bunny_chen joined @GregorVand to talk about open models, production inference, and how RFT is unlocking model customizati…

DGX agent

Our cofounder @the_bunny_chen joined @GregorVand to talk about open models, production inference, and how RFT is unlocking model customization for teams without a dedicated ML org. They covered custom

applicationsfireworks-ai--x
29 Apr 2026
Model Releases

Qwen 3.5 from @Alibaba_Qwen is now available on @FireworksAI_HQ Training Platform across the Managed and Training API workflows. Try SFT, DP…

DGX agent

Qwen 3.5 from @Alibaba_Qwen is now available on @FireworksAI_HQ Training Platform across the Managed and Training API workflows. Try SFT, DPO, RL with smart defaults or your own custom loss function w

model-releasesfireworks-ai--x
29 Apr 2026
Applications

2026: the year of explosive AI innovation. @FireworksAI_HQ Co-founder & CEO Lin Qiao speaks on how the massive expansion of customized infer…

DGX agent

2026: the year of explosive AI innovation. @FireworksAI_HQ Co-founder & CEO Lin Qiao speaks on how the massive expansion of customized inference and models is creating speed of light production to sca

applicationsfireworks-ai--x
28 Apr 2026
Model Releases

DeepSeek V4-Pro on Fireworks. Zoooooom.

DGX agent

DeepSeek V4-Pro is now available on the Fireworks AI platform, likely offering accelerated inference speeds ('Zoooooom' suggests performance optimization). This deployment enables users to access Deep

model-releasesfireworks-ai--x
28 Apr 2026
Model Releases

Excited to support @NVIDIA Nemotron 3 Nano Omni, now available on Fireworks. It's the first open model that handles vision, audio, video, an…

DGX agent

Excited to support @NVIDIA Nemotron 3 Nano Omni, now available on Fireworks. It's the first open model that handles vision, audio, video, and text in a single inference loop. Built for multimodal sub-

model-releasesfireworks-ai--x
28 Apr 2026
Model Releases

GLM 5.1 from @Zai_org is now available on @FireworksAI_HQ Training Platform across the Managed and Training API workflows. Try SFT and DPO w…

DGX agent

GLM 5.1 from @Zai_org is now available on @FireworksAI_HQ Training Platform across the Managed and Training API workflows. Try SFT and DPO with smart defaults or your own custom loss function with a 2

model-releasesfireworks-ai--x
28 Apr 2026
Tools

Prevent prompt injection. safe_tokenization: true Keep your system yours. https://fireworks.ai/blog/safe-tokenization-preventing-prompt-inje…

DGX agent

Safe tokenization is a security feature that helps prevent prompt injection attacks by ensuring that user inputs are properly processed and isolated from system instructions. Fireworks AI discusses ho

toolsfireworks-ai--x
28 Apr 2026
← Previous
1234
Next →