Hany Farid

"All you need is about a minute to two minutes of a person’s voice. There are services that you can pay $5 per month for [that let you] upload your reference audio and clone the voice. Then you can type and get convincing audio in a few seconds. This is text-to-speech.

"There’s also a second way to do this called speech-to-speech. I record a person and clone their voice. And then I record myself saying what I want them to say with all the intonation—bad words and all—and it converts my voice into their voice. It’s all the same underlying generative AI technology.

"For either method, anybody can do this. There is no barrier to entry or technical skill involved."

Comments

Popular posts from this blog

Supporting Artistes (SAs)

Hamza Chaudhry

Injection