Model Releases

Vellium v1.1.0 — Live voice, local STT/TTS and easier llama.cpp setup

Vellium is an open-source, local-first desktop app for AI chat, character roleplay and long-form writing. Recent updates have focused mostly on making local voice and model setups easier to use. Live

DGX agentreddit
model-releasesr-localllama

Vellium is an open-source, local-first desktop app for AI chat, character roleplay and long-form writing. Recent updates have focused mostly on making local voice and model setups easier to use. Live mode now supports microphone input, local or Whisper-compatible speech recognition, streaming TTS, attachments, screen context and the usual chat tools—all inside the same voice interface. Local speech can be installed and configured directly in the app. Whisper Large v3 Turbo Q5_0 is available for recognition, while TeraTTSv2 provides English and Russian voices with realtime playback. The TTS process stays active between responses, avoiding a full model reload for every reply. The llama.cpp setup has also been simplified. Vellium can detect existing llama-server installations, GGUF models and running local endpoints, then configure them as a managed backend. There have been plenty of smaller fixes as well: more reliable TTS streaming, safer runtime archive extraction, better endpoint discovery, improved timeout handling, system certificate support and easier settings navigation. Chats, characters, LoreBooks, writing projects and knowledge collections are stored locally in SQLite. Vellium runs on macOS, Windows and Linux and supports OpenAI-compatible APIs, OpenRouter, LM Studio, Ollama and KoboldCpp. GitHub: https://github.com/tg-prplx/vellium Feedback from people using local voice or roleplay setups would be especially useful-particularly about anything that still feels awkward or unnecessarily complicated. submitted by /u/Possible_Statement84 [link] [comments]

Related

Source: r/LocalLLaMA | 2026-09-01

Loading related sources…