Google launched new Gemini 3.5 audio transcription models, including 3.5 Transcribe supporting 85 languages and auto-removing speech disfluencies, enhancing voice assistant accuracy and opening opportunities for MENA enterprises to automate meetings and voice services.

1 min read

Google's Gemini 3.5 Transcribe: New AI Model for Audio Transcription Supports 85 Languages and Removes 'Ums' and 'Ahs'

Launch of New Gemini 3.5 Models

FAQ

What is Gemini 3.5 Transcribe?

It's a new Google model dedicated to audio transcription, supporting over 85 languages, automatically removing speech disfluencies, and understanding specialized jargon, making it ideal for meetings and audio content.

How does Gemini 3.5 Transcribe compare to previous models?

It offers higher accuracy in noisy environments, better understanding of informal speech, and broader language support, compared to earlier models that struggled with background noise or interrupted speech.

Can MENA enterprises use these models?

Yes, especially with Arabic and many regional languages supported, enterprises can automate meeting transcription, improve voice customer service, and create accurate text content from audio.

When will Gemini 3.5 Pro be available?

Google promised a June release, but no specific date has been announced yet, leaving users waiting for the more advanced model.

Source: The Verge AI

AI-assisted content, human-reviewed.