We all ramble, backtrack, and stumble over our words when we speak out loud. Google‘s new Gemini 3.5 Transcribe is built keeping our imperfect speech in mind. Announced as the company’s most precise speech-to-text model yet, it takes your messy, unpolished speech and turns it into clean, formatted text without expecting you to sound like a news anchor first.
It arrives alongside two companion models, Gemini 3.5 Live and Gemini 3.5 Live Experimental, and together the three make up what Google is branding “Gemini Audio.”