Skip to content
Back to Guavy Wire
Stocks

Google Unveils Most Powerful Speech-to-Text Model Yet

Instruments
GOOGL
Share

Google has released its most powerful speech-to-text model yet, Gemini 3.5 Transcribe, which can automatically process filler words, verbal slips, and repeated expressions during transcription.

The model is designed to convert original speech into text closer to the finished draft, reducing users' subsequent workload of organizing meeting minutes, interview shorthand, and call content.

Gemini 3.5 Transcribe supports over 85 languages and regional variants, can automatically determine the language currently in use, and can continue to transcribe even if the language is switched in one sentence or the same paragraph of dialogue.

The model also provides custom vocabulary lists, which developers can use to guide the model to prioritize recognizing specific contents. Google stated that this feature is usually more effective when the custom vocabulary list is controlled within 100 words.

More on Stocks

Disclaimer: Guavy is a data and market intelligence provider, not an investment adviser. The information, signals, and market analysis provided by the Guavy API and related services are for informational purposes only and are not intended as financial advice, investment recommendations, or an endorsement of any particular trading strategy. Trading in volatile markets, including cryptocurrency, carries significant risk and may not be suitable for all investors. Past performance is not indicative of future results. Users should consult with a qualified financial professional before making any investment decisions. Guavy makes no guarantee of trading profits or financial returns.

Market sentiment intelligence for apps, funds & agents

Location

729 55 Ave SW
Calgary AB T2V 0G4
Canada

© 2026 Guavy Inc