Close Menu
Versa AI hub
  • AI Ethics
  • AI Legislation
  • Business
  • Cybersecurity
  • Media and Entertainment
  • Content Creation
  • Art Generation
  • Research
  • Tools
  • Resources

Subscribe to Updates

Subscribe to our newsletter and stay updated with the latest news and exclusive offers.

What's Hot

Introducing Lyria 3.5 to Google Flow Music

July 29, 2026

How AI is shortening drug discovery timelines in China

July 28, 2026

Introducing real-time generative simulation to surgical robotics

July 27, 2026
Facebook X (Twitter) Instagram
Versa AI hubVersa AI hub
Tuesday, August 18
Facebook X (Twitter) Instagram
Login
  • AI Ethics
  • AI Legislation
  • Business
  • Cybersecurity
  • Media and Entertainment
  • Content Creation
  • Art Generation
  • Research
  • Tools
  • Resources
Versa AI hub
Home»Tools»Gemini 2.5 update from Google Deepmind
Tools

Gemini 2.5 update from Google Deepmind

versatileaiBy versatileaiMay 21, 2025No Comments1 Min Read
Share Facebook Twitter Pinterest LinkedIn Tumblr Reddit Telegram Email
#image_title
Share
Facebook Twitter LinkedIn Pinterest Email

New Gemini 2.5 Features

Native audio output and live API improvements

Today, Live APIs introduce preview versions of audiovisual input and native audio out dialogs, allowing you to directly build conversational experiences with more natural and expressive Gemini.

It also allows users to manipulate tones, accents and speech styles. For example, you can instruct your model to use dramatic voices when telling stories. It also supports the use of the tool and allows you to search for it on your behalf.

You can try out a set of early features including:

An emotional dialogue in which the model detects and responds appropriately to the user’s voice emotions. In ProActual Audio, models can ignore background conversations and know when to respond. The idea in the live API utilizes Gemini’s thinking capabilities to help the model support more complex tasks.

We are also releasing new previews of text-to-speech in 2.5 Pro and 2.5 Flash. These have initial support for multiple speakers, allowing speech from text using two voices via native audio out.

Like native audio dialogs, text-to-speech is expressive and can capture very subtle nuances such as whispers. Works in over 24 languages ​​and seamlessly switch between them.

author avatar
versatileai
See Full Bio
Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
Previous ArticleGoogle’s AI vision is clouded by business model hallucinations • Register
Next Article VEO – Google Deep Mind
versatileai

Related Posts

Tools

Introducing Lyria 3.5 to Google Flow Music

July 29, 2026
Tools

How AI is shortening drug discovery timelines in China

July 28, 2026
Tools

Introducing real-time generative simulation to surgical robotics

July 27, 2026
Add A Comment

Comments are closed.

Top Posts

Humanoid robot performance wins Nvidia Tech boost

August 28, 20255 Views

Agentic AI scaling requires new memory architecture

January 7, 20264 Views

Spread your wings: Introducing the Falcon 180B

October 19, 20254 Views
Stay In Touch
  • YouTube
  • TikTok
  • Twitter
  • Instagram
  • Threads
Latest Reviews

Subscribe to Updates

Subscribe to our newsletter and stay updated with the latest news and exclusive offers.

Most Popular

Humanoid robot performance wins Nvidia Tech boost

August 28, 20255 Views

Agentic AI scaling requires new memory architecture

January 7, 20264 Views

Spread your wings: Introducing the Falcon 180B

October 19, 20254 Views
Don't Miss

Introducing Lyria 3.5 to Google Flow Music

July 29, 2026

How AI is shortening drug discovery timelines in China

July 28, 2026

Introducing real-time generative simulation to surgical robotics

July 27, 2026
Service Area
X (Twitter) Instagram YouTube TikTok Threads RSS
  • About Us
  • Contact Us
  • Privacy Policy
  • Terms and Conditions
  • Disclaimer
© 2026 Versa AI Hub. All Rights Reserved.

Type above and press Enter to search. Press Esc to cancel.

Sign In or Register

Welcome Back!

Login to your account below.

Lost password?