DeepMind has launched its latest audio model, Gemini 3.1, which introduces innovative granular audio tags. These tags provide users with enhanced control over AI-generated speech, allowing for more expressive audio outputs. This advancement signifies a step forward in the capabilities of AI speech synthesis, enabling developers to create more nuanced and engaging audio experiences. As the demand for high-quality AI-generated content continues to rise, Gemini 3.1 positions itself as a pivotal tool for industries reliant on voice technology, from entertainment to customer service. The introduction of this model highlights DeepMind's commitment to pushing the boundaries of artificial intelligence and its applications in real-world scenarios.
Source: DeepMind