In just a few seconds, you can create a copy of a voice—a new simplified model for speech synthesis has emerged…
In just a few seconds, you can create a copy of a voice—a new simplified model for speech synthesis has emerged, capable of reproducing someone else's voice from a short audio clip. It's all very simple: just provide the neural network with a few seconds of a person's recording, and it can speak any text in that same voice. The result sounds quite natural: the quality reaches 48 kHz, which is comparable to a regular audio recording. The most astonishing part is the speed. The model generates speech 150 times faster than real-time playback. In simpler terms, a minute-long text will be voiced in a fraction of a second. At the same time, the artificial intelligence requires less than 1 GB of video memory, so it can be run locally even on a standard PC or laptop. You can download it here (https://github.com/ysharma3501/LuxTTS).
Related posts
- ACE-Step 1.5 XL has been released, an open-source audio model that outperforms…April 13, 2026
- A powerful open-source speech generator has been released — the Tada model was…March 18, 2026
- Alibaba has unveiled its new flagship neural network, Qwen 3.8 Max, with 2.4…August 03, 2026
- Chinese developers have unveiled GLM 5.2 — a new open-source model already…July 12, 2026