AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “together-ai-blog”

GridTimelineEvolution
29 results
CompaniesToolsTechniques

Each lane shows up to 8 recent matching entries, ordered from earlier to later. Tracks load separately to keep the 75,000+ entry wiki fast.

Techniques

TechniqueAgents5 recent entries
13 Apr 2026EinsteinArena: Harnessing the collective intelligence of agents in the wild to advance science

EinsteinArena is a platform where AI agents collaborate and compete on open math problems. AI agents on EinsteinArena have already set 11 new state-of-the-art results on open math problems — including

→24 Apr 2026Accelerate RL rollouts by up to 50% with distribution-aware speculative decoding

This article describes a technique for accelerating reinforcement learning (RL) rollouts using distribution-aware speculative decoding, which can achieve up to 50% speedup improvements. The method lik

HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
→29 Apr 2026DeepSeek-V4 Pro now available on Together AI

DeepSeek-V4 Pro is now available on Together AI with 512K context, controllable reasoning modes, and cached-input pricing for long-context reasoning workloads like code agents, document intelligence,

→19 May 2026Benchmarking inference at scale: coding agents

This article presents benchmarking results for AI coding agents evaluated at scale, likely comparing performance metrics such as code generation accuracy, execution success rates, and inference effici

→29 Jul 2026ThunderAgent: 2x Faster Agentic Inference for Synthetic Data Generation at Scale

ThunderAgent is a program-aware scheduler for agentic inference. By treating each agent workflow as a schedulable program, it eliminates KV cache thrashing to deliver more than 2x single-node throughp

TechniqueFine-tuning1 recent entries
23 Jul 2026The production platform for open-weight AI inference

OpenAI has updated its inference platform to give users full control over performance, cost, and quality without building their own stack—models go live in minutes and support multiple deployments beh

TechniqueMultimodal4 recent entries
24 Apr 2026Accelerate RL rollouts by up to 50% with distribution-aware speculative decoding

This article describes a technique for accelerating reinforcement learning (RL) rollouts using distribution-aware speculative decoding, which can achieve up to 50% speedup improvements. The method lik

→28 Apr 2026Together AI Brings NVIDIA Nemotron 3 Nano Omni to Developers on Day 0

Together AI announced immediate availability of NVIDIA's Nemotron 3 Nano Omni model to developers through its platform on the day of its release. The Nemotron 3 Nano Omni is a lightweight multimodal m

→2 Jun 2026Serving MiniMax-M3 for efficient inference: Unlocking 1M-Token Context and Multimodality Without Regrets

MiniMax-M3 is a large language model capable of handling 1 million token contexts and multimodal inputs while maintaining efficient inference performance. Together AI's blog post discusses techniques

→29 Jul 2026Configuring Dedicated Model Inference

The Together AI platform’s dedicated inference architecture consists of three immutable entities: **configs** (engine, GPU type/count, parallelism and optimization profile), **deployments** (a specifi