Exhibitor login
Bright 24 September 2026

Gemini clones voices with 30 seconds of audio

Gemini clones voices with 30 seconds of audio

Google has developed a new technology called Gemini 3.8 Flash TTS that allows a voice to be replicated based on just 30 seconds of audio recordings. This advance in speech synthesis offers a wide range of applications, such as in podcasts, audiobooks, games, and AI assistants. The technology can significantly contribute to making content creation more efficient and improving user experiences.

The ability to fully clone voices with minimal input can not only accelerate the production of audio content but also enhance its accessibility. This allows creators and companies to easily utilize unique voice profiles without long recording processes. This can be particularly advantageous in the entertainment industry and when developing interactive experiences.

With this innovative step from Google, the question arises of how this technology can be used for both positive and negative purposes. It highlights the need for regulations and guidelines to prevent misuse of voice replication. As the technology becomes more accessible, it is crucial that users are aware of the ethical and legal implications of voice cloning.

The development of Gemini 3.8 demonstrates the ongoing evolution of artificial intelligence and speech technology. Innovations like this harness the power of AI to create personalized experiences and improve interaction between humans and machines. However, it also raises questions about privacy, identity, and the authenticity of digital content.

Read the full article from Bright.