Google says its latest Gemini transcription model can turn your ramblings into structured text

Engadget 

Google has revealed a new AI audio model that it says offers improved speech recognition and transcription. Gemini 3.5 Transcribe is joining Gemini 3.5 Live and Gemini 3.5 Live Experimental in the Gemini Audio family. The company says Gemini 3.5 Transcribe is more adept at speech-to-text than earlier models, with greater precision and the ability to automatically detect more than 85 languages. It says the model can adapt a user's unstructured speech into formatted text. You can use it to make edits with voice commands, and it removes filler words from transcribed speech too. Instead of typing out every thought, Gemini 3.5 Transcribe works with the way you actually talk: Seamlessly handles self-corrections Removes filler words to deliver clean, formatted text Understands your natural intent and speaking style Accurately captures audio in... pic.twitter.com/17YcsDi7Q0