Introducing Gemini 3.5 Transcribe, our new speech to text model with smart transcription, function calling, more precise transcription (lower WER), custom vocabulary support, multi-speaker identification, and support for over 85 languages!
Also with realtime streaming support!