AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
Human
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “together-ai--x”

GridTimelineEvolution
263 results
CompaniesToolsTechniques

Each lane shows up to 8 recent matching entries, ordered from earlier to later. Tracks load separately to keep the 75,000+ entry wiki fast.

Tools

ToolClaude Code7 recent entries
8 Apr 2026GLM-5.1 gives teams a stronger model for coding, tool use, and sustained agent performance on Together AI. Learn more: http://www.together.a…

GLM-5.1 is Z.ai's post-training upgrade to GLM-5, now available on Together AI, delivering a 28% coding performance improvement through a refined reinforcement learning pipeline while retaining the...

→12 Apr 2026On software engineering: → 56.22% SWE-Pro, matching GPT-5.3-Codex → 55.6% VIBE-Pro — full-stack project delivery → SWE Multilingual: 76.5 Fo…

On software engineering: → 56.22% SWE-Pro, matching GPT-5.3-Codex → 55.6% VIBE-Pro — full-stack project delivery → SWE Multilingual: 76.5 For agentic work: → Multi-agent collaboration built into the m

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
→19 May 2026'One thing that we've been seeing recently is that inference benchmarks don't really match production workloads that well.' - @realDanFu, VP…

'One thing that we've been seeing recently is that inference benchmarks don't really match production workloads that well.' - @realDanFu, VP of Kernels When you're running dozens of concurrent coding

→28 Jun 2026More reason why we’re excited about GLM-5.2 on Together 👇 Strong enough for serious coding work, cheap enough to change routing decisions, …

More reason why we’re excited about GLM-5.2 on Together 👇 Strong enough for serious coding work, cheap enough to change routing decisions, and easy to access through the tools developers already use.

→24 Jul 2026.@Kimi_Moonshot K3 lands on Together on Monday! We ran 452 DeepSWE rollouts against Claude Fable 5: near-flagship coding at ~35% of the pric…

.@Kimi_Moonshot K3 lands on Together on Monday! We ran 452 DeepSWE rollouts against Claude Fable 5: near-flagship coding at ~35% of the price, and K3 pulls ahead at higher pass@k's. Full deep-dive: ht

→6 Aug 2026Open models give teams more room to run agent loops, with API prices at a fraction of GPT-5.6 Sol and Claude Fable 5. That matters as planni…

Open models give teams more room to run agent loops, with API prices at a fraction of GPT-5.6 Sol and Claude Fable 5. That matters as planning, tool calls, retries, and long contexts compound token us

→10 Aug 2026Frontier performance you can actually own. Proud to help power DeepSeek-V4-Flash on Ollama's cloud, with the fastest hosted performance avai…

Frontier performance you can actually own. Proud to help power DeepSeek-V4-Flash on Ollama's cloud, with the fastest hosted performance available. Open weights, zero data retention, U.S. & EU hosting.

ToolCursor6 recent entries
18 May 2026Congrats to the @cursor_ai team on Composer 2.5 — a huge milestone for agentic coding models. Together AI, the AI Native Cloud, is proud to …

Congrats to the @cursor_ai team on Composer 2.5 — a huge milestone for agentic coding models. Together AI, the AI Native Cloud, is proud to partner on this launch. Composer 2.5 is pushing the frontier

→10 Jun 2026Learn how @cursor_ai partnered with Together AI to deliver real-time inference for AI-powered coding in this article from @ce_zhang and @rea…

Learn how @cursor_ai partnered with Together AI to deliver real-time inference for AI-powered coding in this article from @ce_zhang and @realDanFu. Cursor's in-editor agents generate code while develo

→2 Jul 202630 billion tokens a month to 400 trillion in a year. That's @cursor_ai, @DecagonAI, @cartesia and hundreds of other teams choosing open infr…

Together AI highlights the rapid growth of AI token consumption across multiple companies and teams, noting an increase from 30 billion tokens monthly to 400 trillion annually, with platforms like Cur

→7 Jul 2026This is what latency optimization looks like below the API 👇 Together ATLAS, NVIDIA Blackwell, CUDA, TensorRT-LLM, Dynamo, and custom kerne…

This is what latency optimization looks like below the API 👇 Together ATLAS, NVIDIA Blackwell, CUDA, TensorRT-LLM, Dynamo, and custom kernels all working together to make inference faster for users. P

→20 Jul 2026YC and Together AI are partnering to bring the first dedicated YC GPU cluster online, giving YC startups easier access to the compute they n…

YC and Together AI are partnering to bring the first dedicated YC GPU cluster online, giving YC startups easier access to the compute they need to build and scale. In this Founder Fireside, YC's @agup

→6 Aug 2026Congrats to @mattrubens and the Roomote team on the launch. Builders can use Together AI as an inference provider in Roomote and assign diff…

Congrats to @mattrubens and the Roomote team on the launch. Builders can use Together AI as an inference provider in Roomote and assign different open models to coding, planning, vision, and review ac

ToolOllama4 recent entries
22 Apr 2026Introducing Kimi K2.6 from @Kimi_Moonshot, a multimodal agentic model with Agent Swarm scaling to 300 sub-agents and long-horizon coding sta…

Introducing Kimi K2.6 from @Kimi_Moonshot, a multimodal agentic model with Agent Swarm scaling to 300 sub-agents and long-horizon coding stability. AI natives can now use Kimi K2.6 on Together AI and

→9 Jul 2026Congrats to @ollama on the fundraise and 9M+ active builders.🚀 Open models are becoming the default path for developers who want to build, …

Congrats to @ollama on the fundraise and 9M+ active builders.🚀 Open models are becoming the default path for developers who want to build, run, and own their AI stack. We're excited to support the eco

→10 Aug 2026Frontier performance you can actually own. Proud to help power DeepSeek-V4-Flash on Ollama's cloud, with the fastest hosted performance avai…

Frontier performance you can actually own. Proud to help power DeepSeek-V4-Flash on Ollama's cloud, with the fastest hosted performance available. Open weights, zero data retention, U.S. & EU hosting.

→11 Aug 2026DeepSeek V4 Flash 0731 is now available to fine-tune on Together AI. Specialize it for coding, tool use, and your own domain with SFT or DPO…

DeepSeek V4 Flash 0731 is now available to fine-tune on Together AI. Specialize it for coding, tool use, and your own domain with SFT or DPO, then deploy the fine-tuned model on Together AI for produc

ToolVercel AI5 recent entries
9 Apr 2026Gemma 4 31B brings dense multimodal reasoning to Together AI. Try Now: http://www.together.ai/models/gemma-4-31b

Google's Gemma 4 31B is a dense multimodal model from Google DeepMind now available on Together AI's serverless infrastructure via the endpoint `google/gemma-4-31B-it`. It features a 256K context ...

→12 Apr 2026Run it on the AI Native Cloud — serverless and dedicated infrastructure. https://www.together.ai/models/minimax-m2-7

MiniMax-M2 is a large-scale mixture-of-experts (MoE) language model available for inference on Together AI's platform, accessible via both serverless and dedicated infrastructure options. The model ca

→12 Apr 2026MiniMax M2.7 is now on Together AI. Trained by letting it run its own RL loop, resulting in the highest open-source score on MLE Bench Lite.

MiniMax M2.7 is now available on Together AI, offering a powerful open-source model trained using a self-directed reinforcement learning loop. This training methodology enabled the model to achieve th

→20 Apr 2026Our researchers are heading to ICLR with new work: model efficiency, long-context reasoning, next-gen attention and decoding, and more. Chec…

Our researchers are heading to ICLR with new work: model efficiency, long-context reasoning, next-gen attention and decoding, and more. Check out what we've been building 👇🏼 #TogetherResearch #AINativ

→22 Apr 2026Try Kimi K2.6 now on the AI Native Cloud: http://www.together.ai/models/kimi-k26#

Kimi K2.6 is now available for use on Together AI's cloud platform, which offers AI model deployment and inference services. This announcement indicates the model has been added to Together AI's roste

ToolHugging Face8 recent entries
16 Apr 2026We’re back 🔥. Thrilled to be named once again to the @Forbes AI 50. The AI Native Cloud, built for the full AI lifecycle: fast inference, o…

We’re back 🔥. Thrilled to be named once again to the @Forbes AI 50. The AI Native Cloud, built for the full AI lifecycle: fast inference, open models, fine-tuning at scale. https://www.forbes.com/list

→20 Apr 2026Our researchers are heading to ICLR with new work: model efficiency, long-context reasoning, next-gen attention and decoding, and more. Chec…

Our researchers are heading to ICLR with new work: model efficiency, long-context reasoning, next-gen attention and decoding, and more. Check out what we've been building 👇🏼 #TogetherResearch #AINativ

→23 Apr 202630B -> 300T tokens per month YoY @togethercompute

Together AI achieved a 10,000x increase in monthly token throughput over one year, scaling from 30 billion to 300 trillion tokens per month. This milestone demonstrates significant growth in their API

→8 May 2026The gap between 'this looks cool' and 'I'm actually running it' used to be a day or two of setup. Not anymore. Deploy any Hugging Face model…

The gap between 'this looks cool' and 'I'm actually running it' used to be a day or two of setup. Not anymore. Deploy any Hugging Face model on the AI Native Cloud in a single session 👇🏼 https://www.t

→24 Jun 2026Read the full story: https://www.theinformation.com/newsletters/applied-ai/open-source-growth-boosts-together-ai-hugging-face

Together AI highlights how the growth of open-source AI models is benefiting their platform and the broader ecosystem. The article likely discusses how open-source initiatives, including collaboration

→26 Jun 2026As token usage explodes, model choice becomes product strategy. Teams are already testing models like GLM-5.2 because they want frontier qua…

As token usage explodes, model choice becomes product strategy. Teams are already testing models like GLM-5.2 because they want frontier quality, better tokenomics, and more control over cost, data, a

→27 Jul 2026Kimi K3 is now available directly from its Hugging Face model page through Together AI, give it a try!

Kimi K3 is now available directly from its Hugging Face model page through Together AI, give it a try! Kimi K3 in @huggingface Inference Providers is live via @togethercompute 3/M input tokens, 15/M o

→5 Aug 2026.@Kimi_Moonshot benchmarked K3 endpoints across major inference providers. Together AI leads or ties for #1 on 3 of 4 benchmarks: OCRBench, …

.@Kimi_Moonshot benchmarked K3 endpoints across major inference providers. Together AI leads or ties for #1 on 3 of 4 benchmarks: OCRBench, MMMU Pro Vision, and DeepSWE. Open models like Kimi K3 have