| ▲ | kingstnap 5 hours ago | |||||||||||||
Yesterday night I was doing a project with QwenTTS 1.7B. After some debugging, making a clean dataset with clean recordings, and experimenting with a good fine tune recipe (much props to the new GPT models yesterday being cheaper). I was able to make a robo-me that sounds absurdly good, family was shocked, all in a matter of a few hours. So yeah the cat is out of the bag for sure. | ||||||||||||||
| ▲ | MacNCheese23 3 hours ago | parent | next [-] | |||||||||||||
Yeah I was doing that at the beginning of this year with voice samples locally from hollywood-stars with Qwen3-TTS. It took like 1min - capture something from a youtube or video and put in your own text. It worked also really good for a german test. Made a voice message for my wife from one of our favorite actors, telling here how nice it would be to make some breakfast :D | ||||||||||||||
| ▲ | rpastuszak an hour ago | parent | prev | next [-] | |||||||||||||
Any chance you could share a bit more detail? I’d love to try this myself but could use some proven structure / approach. | ||||||||||||||
| ||||||||||||||
| ▲ | yieldcrv 4 hours ago | parent | prev | next [-] | |||||||||||||
Its so crazy to me how prevalent bad AI voices are, when local models can do such good AI voices | ||||||||||||||
| ||||||||||||||
| ▲ | 3 hours ago | parent | prev [-] | |||||||||||||
| [deleted] | ||||||||||||||