I have discovered some Machine Learning Models that allow you to create a realistic voice for text. I'm planning to build an application that can convert the text to speech and clone your voice to talk in a realistic way. I'm not sure if there will be an audience for this, as it would be a much cheaper alternative to ElevenLabs and can handle longer sequences. I'm looking for feedback on this idea.
Eleven Labs: https://beta.elevenlabs.io/