This has been my dream since I was a teen - and now the tech is real. This is a fork of an existing speech-to-speech project by Huggingface, with a lot of additions to add the right chimes, the voice* and the overall wakeword architecture.

I’ve had it running for a day and … I like it. I’ll most likely wire up a version of this throughout the house so the whole family can live the trekkie life :)

*) Cloning Barrett’s voice … I know. It’s just little old me showing that we now have the tech, not some commercial venture. It’s done with the utmost of respect for her craft throughout all those years!

  • troed@fedia.ioOP
    link
    fedilink
    arrow-up
    3
    ·
    2 days ago

    Qwen3-TTS does excellent cloning and is very fast. On my workstation (5060Ti) it renders the audio at around 2.5x realtime. I’ve added an example with my parameters now.