Model Releases
🇺🇸 xAI just dropped a real-time speech-to-text model built for voice apps: high limits, multi-language support, the works. Priced at just …
🇺🇸 xAI just dropped a real-time speech-to-text model built for voice apps: high limits, multi-language support, the works. Priced at just 0.10–0.20/hour, crushing most competitors. Quiet release, loud
🇺🇸 xAI just dropped a real-time speech-to-text model built for voice apps: high limits, multi-language support, the works. Priced at just 0.10–0.20/hour, crushing most competitors. Quiet release, loud warning to the rest." Source: @XFreeze, @xAI, @elonmusk Grok took #1 in healthcare. Not by a little. 300K+ votes. 308 models. Real people comparing real answers. Beating Claude. Beating Gemini. By the numbers. People getting actual answers that help. That’s the whole point. @xAI @Grok
Related
- MTR-DuplexBench: Towards a Comprehensive Evaluation of Multi-Round Conversations for Full-Duplex Speech Language Models
- SpeechMedAssist: Efficiently and Effectively Adapting Speech Language Models for Medical Consultation
- NIM4-ASR: Towards Efficient, Robust, and Customizable Real-Time LLM-Based ASR
- DeepL, best known for its text translation tools, launches DeepL Voice-to-Voice, which enables real-time spoken translation, with add-ons for services like Zoom (Ivan Mehta/TechCrunch)
- Google rolls out Gemini 3.1 Flash TTS, a text-to-speech model with support for over 70 languages and audio tags that give developers granular speech control (Matthias Bastian/The Decoder)
Source: Elon Musk (X) | 2026-04-17