TechniqueAgents5 recent entries13 Apr 2026EinsteinArena: Harnessing the collective intelligence of agents in the wild to advance scienceEinsteinArena is a platform where AI agents collaborate and compete on open math problems. AI agents on EinsteinArena have already set 11 new state-of-the-art results on open math problems — including→24 Apr 2026Accelerate RL rollouts by up to 50% with distribution-aware speculative decodingThis article describes a technique for accelerating reinforcement learning (RL) rollouts using distribution-aware speculative decoding, which can achieve up to 50% speedup improvements. The method lik
TechniqueFine-tuning1 recent entries23 Jul 2026The production platform for open-weight AI inferenceOpenAI has updated its inference platform to give users full control over performance, cost, and quality without building their own stack—models go live in minutes and support multiple deployments beh
TechniqueMultimodal4 recent entries24 Apr 2026Accelerate RL rollouts by up to 50% with distribution-aware speculative decodingThis article describes a technique for accelerating reinforcement learning (RL) rollouts using distribution-aware speculative decoding, which can achieve up to 50% speedup improvements. The method lik→28 Apr 2026Together AI Brings NVIDIA Nemotron 3 Nano Omni to Developers on Day 0Together AI announced immediate availability of NVIDIA's Nemotron 3 Nano Omni model to developers through its platform on the day of its release. The Nemotron 3 Nano Omni is a lightweight multimodal m→2 Jun 2026Serving MiniMax-M3 for efficient inference: Unlocking 1M-Token Context and Multimodality Without RegretsMiniMax-M3 is a large language model capable of handling 1 million token contexts and multimodal inputs while maintaining efficient inference performance. Together AI's blog post discusses techniques →29 Jul 2026Configuring Dedicated Model InferenceThe Together AI platform’s dedicated inference architecture consists of three immutable entities: **configs** (engine, GPU type/count, parallelism and optimization profile), **deployments** (a specifi