Applications
Excited to speak at PyCon US @pycon about production LLM inference systems, runtime optimization, and next-gen engine design — including som…
Excited to speak at PyCon US @pycon about production LLM inference systems, runtime optimization, and next-gen engine design — including some of the new ideas behind the recently open-sourced project
Excited to speak at PyCon US @pycon about production LLM inference systems, runtime optimization, and next-gen engine design — including some of the new ideas behind the recently open-sourced project TokenSpeed @lightseekorg See you there! @zhyncs42, Sr. Director of Inference, is taking the stage at PyCon US on May 16! He'll walk through the real work behind running LLM inference in production. His talk covers: → What role Python actually plays in inference runtime optimization → The challenges seen in real deploym…
Source: Together AI (X) | 2026-05-08