August 2026
Google's latest speech model joins Whisper Memos, with automatic language detection, speaker separation in Meeting mode, and no upgrade required.
Whisper Memos now supports Gemini 3.5 Transcribe, Google's latest speech recognition model. It detects the spoken language on its own, handles recordings up to one hour, and is available on every plan — no upgrade needed.
It's a strong all-rounder: accurate on everyday recordings, comfortable with mixed languages, and fast enough that you rarely wait for a transcript.
Gemini also knows who is speaking. If you pick it as your model, meetings up to 30 minutes go to Gemini and come back split by speaker.
Meetings you don't pick a model for keep using ElevenLabs Scribe, which is still our most dependable option for separating speakers. The same goes for meetings over 30 minutes, and for meetings where you use custom words — Gemini can't combine a custom vocabulary with speaker separation yet.
Open Settings → AI Model and pick Gemini 3.5 Transcribe. Every new recording will use it; existing memos stay as they are. Recordings longer than an hour fall back to the usual models automatically.
Update Whisper Memos to try Gemini 3.5 Transcribe on your next recording.