Model Releases
Vellium v1.1.0 — Live voice, local STT/TTS and easier llama.cpp setup
Vellium is an open-source, local-first desktop app for AI chat, character roleplay and long-form writing. Recent updates have focused mostly on making local voice and model setups easier to use. Live
Vellium is an open-source, local-first desktop app for AI chat, character roleplay and long-form writing. Recent updates have focused mostly on making local voice and model setups easier to use. Live mode now supports microphone input, local or Whisper-compatible speech recognition, streaming TTS, attachments, screen context and the usual chat tools—all inside the same voice interface. Local speech can be installed and configured directly in the app. Whisper Large v3 Turbo Q5_0 is available for recognition, while TeraTTSv2 provides English and Russian voices with realtime playback. The TTS process stays active between responses, avoiding a full model reload for every reply. The llama.cpp setup has also been simplified. Vellium can detect existing llama-server installations, GGUF models and running local endpoints, then configure them as a managed backend. There have been plenty of smaller fixes as well: more reliable TTS streaming, safer runtime archive extraction, better endpoint discovery, improved timeout handling, system certificate support and easier settings navigation. Chats, characters, LoreBooks, writing projects and knowledge collections are stored locally in SQLite. Vellium runs on macOS, Windows and Linux and supports OpenAI-compatible APIs, OpenRouter, LM Studio, Ollama and KoboldCpp. GitHub: https://github.com/tg-prplx/vellium Feedback from people using local voice or roleplay setups would be especially useful-particularly about anything that still feels awkward or unnecessarily complicated. submitted by /u/Possible_Statement84 [link] [comments]
Related
- TontaubeV1 - Open TTS model release for local long-form generation
- Update your chat template for dsv4 if you're using llama.cpp
- Parlor v2: best-effort fully local GPT-Live clone on an M3 Pro
- Introducing Unsloth Desktop app
- How do you deal with long-context sessions after restarting llama.cpp?
Source: r/LocalLLaMA | 2026-09-01