Hardware

Together AI STT models now hold the top two spots for transcription speed on the @ArtificialAnlys Speech to Text leaderboard. NVIDIA Parakee…

Together AI STT models now hold the top two spots for transcription speed on the @ArtificialAnlys Speech to Text leaderboard. NVIDIA Parakeet TDT 0.6B V3 on Together AI ranks #1, transcribing 303 seco

DGX agentx-post
hardwaretogether-ai--x

Together AI STT models now hold the top two spots for transcription speed on the @ArtificialAnlys Speech to Text leaderboard. NVIDIA Parakeet TDT 0.6B V3 on Together AI ranks #1, transcribing 303 seconds of audio per second of processing time. → Fastest STT model measured by Artificial Analysis → $1.50 per 1,000 minutes of audio → 4.6% AA-WER across 3 real-world datasets For AI natives building real-time voice agents, fast STT is core infrastructure. Running leading speech models on the AI Native Cloud gives teams more room to keep latency low across transcription, reasoning, and response. Full leaderboard: https://lnkd.in/gw8jXPNQ

Source: Together AI (X) | 2026-05-14

Loading related sources…