← All stories
● Covered by 6 sources · 11 reportsMedium impact10 neutral

Google may soon allow customization of Gemini voice parameters

🔄 Updated 4d ago — new reporting from Ars Technica
New to BrevFeed? We gather this story from every outlet covering it into one summary — ranked by real-world impact, not just the latest headline — so you never miss what matters. What is BrevFeed? →

Key points

  • Customization includes Energy, Formality, Speed, and Warmth.
  • Users can select low, medium, or high increments for each parameter.
  • Voice settings will apply in Gemini Live and chat experiences.
  • Gboard's Writing tools on Pixel 11 series offer "Personalized suggestions" with Gemini Intelligence.
  • Personalized suggestions are tailored to personal style based on conversation, text input, and typing patterns.
  • A "Suggested" tab provides draft suggestions.
  • Gboard uses typing history and screen context from messaging apps.
  • Private AI Compute processes data encrypted on-device and in the cloud.
  • A new menu allows enabling/disabling "Personalized suggestions."
  • Gboard's new voice-to-text tool is named "Rambler."
  • "Rambler" removes filler words and adds formatting like bullet lists and new paragraphs.
  • "Rambler" lacks a live transcription preview.
  • Nothing's upcoming 'Essential Voice' update demonstrates a live transcription preview.
  • Gemini 3.5 Transcribe is a new AI transcription model.
  • Gemini 3.5 Transcribe follows the launch of 3.5 Live Translate.
  • Gemini 3.5 Transcribe detects over 85 languages.
  • Gemini 3.5 Transcribe is rolling out to macOS Gemini app users.
  • Gemini 3.5 Transcribe is an advancement from Chirp 3.
  • Gemini 3.5 Transcribe improves multilingual performance and accuracy.
  • Gemini 3.5 Transcribe allows users to edit naturally with voice.
  • Gemini 3.5 Transcribe adapts to specialized jargon through custom vocabularies.
  • Gemini 3.5 Transcribe is Google's most precise speech-to-text model yet.
  • Gemini 3.5 Transcribe handles natural speech patterns and background noise.
  • Gemini 3.5 Transcribe converts raw audio directly into accurate, polished, formatted text.
  • Gemini 3.5 Transcribe handles self-corrections.
  • Gemini 3.5 Transcribe achieves an average Word Error Rate (WER) of 4.0% for streaming.
  • Gemini 3.5 Transcribe achieves an average Word Error Rate (WER) of 2.6% for non-streaming.
  • Gemini 3.5 Transcribe accurately captures alphanumeric entities.
  • Gemini 3.5 Transcribe is available through the Gemini API for developers.
  • Gemini 3.5 Transcribe will be coming to Chrome.
  • Gemini 3.5 Transcribe will let you use speech-to-text in any web field in Chrome.

Voice Customization in Gemini

Google is preparing a feature in the latest beta of its app that would allow users to customize the Gemini voice. The new capability focuses on four distinct voice characteristics: Energy, Formality, Speed, and Warmth, which can be adjusted in low, medium, or high values.

Implementation Details

The anticipated customization features have been discovered in the code of the Google app version 17.41.12 beta. Users will expect sliders for each parameter, giving them control over how the Gemini voice interacts with them.

Comparative Features

This potential update comes after iOS 27 introduced similar voice customization options for Siri, which adds a competitive dimension between Apple and Google in voice AI. Customizations in iOS affect both Siri's interaction and applications like Maps and Safari.

Future of Voice Technology

As personalization becomes a significant focus in voice technology, these developments could lead to enhanced user experiences. Google has yet to officially announce this feature, but it reflects a shift towards more adaptable voice AI solutions.

Updates

🕒 2026-08-26 · new reporting from The Verge, 9to5Google, Google DeepMind, Engadget
  • Gemini 3.5 Transcribe is a new AI transcription model.
  • Gemini 3.5 Transcribe follows the launch of 3.5 Live Translate.
  • Gemini 3.5 Transcribe detects over 85 languages.
  • Gemini 3.5 Transcribe is rolling out to macOS Gemini app users.
  • Gemini 3.5 Transcribe is an advancement from Chirp 3.
  • Gemini 3.5 Transcribe improves multilingual performance and accuracy.
  • Gemini 3.5 Transcribe allows users to edit naturally with voice.
  • Gemini 3.5 Transcribe adapts to specialized jargon through custom vocabularies.
  • Gemini 3.5 Transcribe is Google's most precise speech-to-text model yet.
  • Gemini 3.5 Transcribe handles natural speech patterns and background noise.
  • Gemini 3.5 Transcribe converts raw audio directly into accurate, polished, formatted text.
  • Gemini 3.5 Transcribe handles self-corrections.
  • Gemini 3.5 Transcribe achieves an average Word Error Rate (WER) of 4.0% for streaming.
  • Gemini 3.5 Transcribe achieves an average Word Error Rate (WER) of 2.6% for non-streaming.
  • Gemini 3.5 Transcribe accurately captures alphanumeric entities.
  • Gemini 3.5 Transcribe is available through the Gemini API for developers.
  • Gemini 3.5 Transcribe will be coming to Chrome.
  • Gemini 3.5 Transcribe will let you use speech-to-text in any web field in Chrome.
🕒 2026-08-25 · new reporting from 9to5Google
  • Gboard's new voice-to-text tool is named "Rambler."
  • "Rambler" removes filler words and adds formatting like bullet lists and new paragraphs.
  • "Rambler" lacks a live transcription preview.
  • Nothing's upcoming 'Essential Voice' update demonstrates a live transcription preview.
🕒 2026-08-20 · new reporting from 9to5Google
  • Gboard's Writing tools on Pixel 11 series offer "Personalized suggestions" with Gemini Intelligence.
  • Personalized suggestions are tailored to personal style based on conversation, text input, and typing patterns.
  • A "Suggested" tab provides draft suggestions.
  • Gboard uses typing history and screen context from messaging apps.
  • Private AI Compute processes data encrypted on-device and in the cloud.
  • A new menu allows enabling/disabling "Personalized suggestions."

✨ This summary was generated by AI from the outlets' reporting listed below. It is not independently verified and may contain errors — check the original sources. How BrevFeed works →

The daily brief

One email each morning: the day's tech stories, clustered across outlets and summarized. No account needed.

One email a day. Unsubscribe in one click, any time.

Today's brief

Spend a few minutes, get the whole day. Every topic's top stories in one hands-free rundown — listen, watch, or read the transcript.

~7 min · 6 stories · Aug 30

▶ Play today's brief Listen on Spotify

New every morning, and the back catalogue is archived by date.

How outlets covered it

Google has released Gemini 3.5 Transcribe, an AI model designed for speech-to-text conversion that polishes voice input by removing filler words and allowing on-the-fly corrections. This model offers significant speed improvements and reduced error rates compared to its predecessor, Chirp 3, and will be integrated across the Google ecosystem.

Google announced Gemini 3.5 Transcribe, a new AI audio model that offers enhanced speech recognition, automatic language detection, and the ability to convert unstructured speech into formatted text. This development matters because it integrates advanced voice input capabilities across Google products and web environments, potentially changing how users interact with digital interfaces.

Google has launched Gemini 3.5 Transcribe, a new speech-to-text model available through the Gemini API, designed for intelligent voice interactions. This release allows developers to integrate advanced transcription capabilities, previously used in Google's own products, into their applications.

Google released Gemini 3.5 Transcribe, an updated AI transcription model that automatically removes filler words, formats text, and adapts to specialized jargon through custom vocabularies. This update improves multilingual performance and accuracy compared to its predecessor, Chirp 3, and is rolling out to macOS Gemini app users and Android's Rambler dictation feature.

Google introduced Gemini 3.5 Transcribe, a new speech-to-text model that offers improved precision and handles natural speech patterns, background noise, and jargon. This model is already integrated into products like Gboard Rambler and the Gemini macOS app, and will be coming to Chrome, enhancing voice interaction and transcription capabilities across Google's ecosystem.

Google's 'Rambler' voice-to-text feature on the Pixel 11, powered by Gemini Intelligence, effectively processes speech into coherent messages by removing filler words and adding formatting. However, it lacks a live transcription preview, which requires users to wait for full processing before seeing the output, a feature that Nothing's upcoming 'Essential Voice' update demonstrates.

Google has updated Gboard's Writing tools on the Pixel 11 series with Gemini Intelligence to include "Personalized suggestions." This feature provides draft suggestions tailored to a user's personal style based on current conversation, text input, and typing patterns, utilizing on-device and cloud-based Private AI Compute.

Google's new Gboard Rambler feature on the Pixel 11 series processes speech after recording, rather than providing real-time transcription. This allows it to refine text, filter filler words, and enable post-transcription editing and style changes, differing from previous real-time speech-to-text systems.

Google's new Pixel 11 series debuts several AI-powered features, including Gboard Rambler for refined voice transcription, an updated Proactive Assistance (formerly Magic Cue) with expanded contextual suggestions, and a location-aware At a Glance display. These updates integrate Gemini Intelligence to automate tasks and provide more relevant information directly on the device.

Google is rolling out advanced voice control capabilities for Gemini on macOS, including intelligent dictation that refines spoken words and a feature that understands screen context to perform complex tasks. This update allows users to interact with their desktop using natural language for tasks like summarizing documents, rewriting text, and generating images, aiming to provide a speech-to-text experience similar to Gboard Rambler.

Google's upcoming Gemini update could enable users to customize voice characteristics such as Energy, Formality, Speed, and Warmth. This feature will enhance user interaction with Gemini across various applications, likely improving personalization options for voice interactions.