Applications

Excited to speak at PyCon US @pycon about production LLM inference systems, runtime optimization, and next-gen engine design — including som…

Excited to speak at PyCon US @pycon about production LLM inference systems, runtime optimization, and next-gen engine design — including some of the new ideas behind the recently open-sourced project

DGX agentx-post
applicationstogether-ai--x

Excited to speak at PyCon US @pycon about production LLM inference systems, runtime optimization, and next-gen engine design — including some of the new ideas behind the recently open-sourced project TokenSpeed @lightseekorg See you there! @zhyncs42, Sr. Director of Inference, is taking the stage at PyCon US on May 16! He'll walk through the real work behind running LLM inference in production. His talk covers: → What role Python actually plays in inference runtime optimization → The challenges seen in real deploym…

Source: Together AI (X) | 2026-05-08

Loading related sources…