Tutorials
The one thing I still don't know how to do: TTS/singing a specific song but with a specific voice
The specific Reddit post could not be retrieved from the search results. However, based on the context of the URL and related results, I can provide the following best-effort summary based on what ...
The specific Reddit post could not be retrieved from the search results. However, based on the context of the URL and related results, I can provide the following best-effort summary based on what is available:
A user on r/StableDiffusion raises an unsolved workflow challenge: generating a specific song's audio with a specific cloned or custom voice — combining singing voice synthesis (SVS) with voice cloning in a controllable way. While tools like TorToiSe TTS handle speech cloning and projects like DiffSinger address singing voice synthesis, no straightforward, unified local workflow exists that reliably performs both tasks together. The post invites the community to share techniques or tools that can bridge TTS voice cloning and AI singing for a target song and target voice simultaneously.
Related
- https://x.com/ben_golub/status/2042079804271313343
- Similar issue here. Pretty much everyone in the field had similar concerns; anyone with substantial open source software had a mailing list;…
- Speaker room listens in on @badlogicgames
- Building intelligent audio search with Amazon Nova Embeddings: A deep dive into semantic audio understanding
Source: tutorials