# Top 3 new model launches at Gemini Audio at Night

## Executive summary

Google DeepMind introduced three major audio model advancements: Gemini 3.8 TTS for highly personalized voice generation, Gemini 3.5 Live Translate for real-time, multi-language audio/video translation, and enhanced Gemini Live models that integrate multimodal understanding, function calls, and tool calls into developer projects.

## Key takeaways

- Gemini 3.8 TTS Voice Personalization: New text-to-speech models allow for compelling voice personalization, enabling developers to guide the voice to have specific emotions, resonance, and incorporate elements like pauses.
- Gemini 3.5 Live Translation: A new feature enabling real-time translation for both video and audio input feeds across over 100 languages.
- Gemini Live Models Enhancements: Updated Gemini Live models allow developers to integrate multimodal understanding, live interactions, function calls, and tool calls directly into their developer projects.

## Technical details

- Text-to-Speech (TTS): Gemini 3.8 TTS models support voice personalization, allowing control over specific emotions, resonance, and pacing (e.g., incorporating pauses).
- Real-Time Translation: Gemini 3.5 Live Translate processes both video and audio input feeds, translating content into over 100 languages.
- Multimodal Development: Gemini Live models enhance developer capabilities by allowing the integration of multimodal understanding, live interactions, function calls, and tool calls.

## Practical implications

- Developers can build complex, interactive applications by integrating advanced audio features like real-time translation and personalized voice generation.
- The inclusion of function calls and tool calls within Gemini Live models suggests direct integration into developer workflows and project architectures.
- These models facilitate the creation of highly localized and accessible applications supporting global communication.

## Topics

Audio AI, Generative AI, Text-to-Speech, Real-Time Translation, Multimodal AI, Gemini Audio, Google Developers Subscription

Source: https://www.youtube.com/watch?v=lTQHImoeuEY
