Skip to content
Close Menu
CryptoAINews
  • Cryptocurrency
  • Blockchain
  • Bitcoin News
  • Altcoins
  • Crypto Market Trends
  • Crypto Mining
  • Ethereum
  • AI News
  • Sponsored
  • Advertise
Trending
  • 7 Google Workspace with Gemini tools to use this semester
  • Start your ice fishing journey: essential tips for new players
  • The best ice fishing casino apps: seamless gameplay and rewarding features
  • Conoce los juegos destacados en Pinup casino: slots, ruleta y más
  • Aplicația Chicken Road APK: cum să o descarci și să te bucuri de jocuri
  • Anthropic continues compute-gobbling streak in $45 billion deal with Nscale
  • 5 Small Daily Rituals That Can Help You Look Forward to the End of the Day
  • BitMEX Sets Close-Only Risk Limits Ahead Of September Wind-Down
  • AI News
  • Cryptocurrency
  • Blockchain
  • Bitcoin News
  • Altcoins
  • Crypto Market Trends
  • Crypto Mining
  • Ethereum
  • Sponsored
  • Advertise
CryptoAINews
  • Cryptocurrency
  • Blockchain
  • Bitcoin News
  • Altcoins
  • Crypto Market Trends
  • Crypto Mining
  • Ethereum
  • AI News
  • Sponsored
  • Advertise
CryptoAINews
Home » AI News » Introducing Gemini 3.5 Transcribe
gemini 3 5 transcribe.width 1300
AI News

Introducing Gemini 3.5 Transcribe

CryptoAINewsBy CryptoAINewsAugust 26, 2026No Comments2 Mins Read
Share
Facebook Twitter LinkedIn Pinterest Email


In the present day, we’re introducing Gemini 3.5 Transcribe, our most exact speech-to-text mannequin but, designed for clever voice interactions. In contrast to typical speech recognition fashions that battle with background noise, advanced jargon, and disfluency cleanup, Gemini 3.5 Transcribe converts uncooked audio instantly into correct, polished, formatted textual content.

Throughout our merchandise just like the Gemini app and on Android, we’ve seen customers already benefiting from this transcription mannequin with new voice capabilities like Rambler on Android and within the Gemini app on macOS. Now, builders can construct related capabilities with Gemini 3.5 Transcribe within the Gemini API in Google AI Studio and Gemini Enterprise Agent Platform.

We have constructed 3.5 Transcribe to plug seamlessly into your developer workflows, whether or not you’re constructing voice brokers, real-time captioning instruments, or post-call analytics pipelines. The mannequin is obtainable throughout two separate APIs:

  • Actual-time streaming: Delivers steady, bidirectional streaming with sub-second latency for interactive voice apps through the Live API utilizing gemini-3.5-transcribe-live.
  • Pre-recorded audio processing: Transcribes recorded audio, conferences, name logs, and extra with speaker attribution and word-level timestamps through the Interactions API utilizing gemini-3.5-transcribe.

Get extra exact and clever transcription

Gemini 3.5 Transcribe is designed to seize your pure talking model to raised perceive your intent and acknowledge customized vocabulary, so you possibly can execute duties together with your voice.

  • Good transcription: Seamlessly handles self-corrections (like “let’s meet Tuesday—no, Wednesday”), removes filler phrases (“ums” and ‘“ahs”), auto-formats your textual content.
  • Perform calling: The mannequin can delegate advanced duties (reminiscent of picture era and file evaluation) to different Gemini fashions through perform calls. At the moment accessible within the Gemini macOS app.
  • Extra exact transcription: As measured by Synthetic Evaluation, achieves a median Phrase Error Charge (WER) of 4.0% for streaming and a couple of.6% for non-streaming use-cases. It exhibits sturdy efficiency throughout noisy, real-world environments, precisely capturing alphanumeric entities like postal codes and order IDs.
  • Customized vocabulary: Acknowledges specialised jargon and distinctive spellings by seamlessly adapting transcriptions to your supplied customized vocabulary.
  • International language assist: Robotically detects and transcribes over 85 languages, seamlessly dealing with regional accents and numerous dialects.
  • Multi-speaker identification: Precisely attributes speech in pre-recorded audio with timestamps for as much as three audio system (assist for 3+ audio system is experimental).



Source link

Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
CryptoAINews
  • Website

Related Posts

7 Google Workspace with Gemini tools to use this semester

August 27, 2026

Anthropic continues compute-gobbling streak in $45 billion deal with Nscale

August 26, 2026

CISA confirms hackers targeted over 100 US water systems during July

August 26, 2026

AI, IP, and the future of innovation

August 26, 2026
Add A Comment

Comments are closed.

About us

CryptoAINews is an independent digital publication focused on cryptocurrency, blockchain, and artificial intelligence news.

The platform is owned and operated by Robert Grabarevic, providing timely news coverage, market updates, and educational content for a global audience interested in emerging technologies and digital finance.

CryptoAINews is committed to transparent reporting, responsible publishing, and delivering informative content based on publicly available data, verified sources, and industry developments.

All content published on this website is for informational purposes only and does not constitute financial or investment advice.

Top Insights

7 Google Workspace with Gemini tools to use this semester

August 27, 2026

Start your ice fishing journey: essential tips for new players

August 27, 2026

The best ice fishing casino apps: seamless gameplay and rewarding features

August 27, 2026
Categories
  • ! Без рубрики
  • Advertise
  • AI News
  • Altcoins
  • Bitcoin News
  • Blockchain
  • Crypto Market Trends
  • Crypto Mining
  • Cryptocurrency
  • Ethereum
  • Live Casino Bet
  • public
  • Sponsored
  • Imprint-Legal-Notice
  • Author / Publisher Bio
  • Privacy Policy
© 2025 CryptoAINews – Owned & Operated by Robert Grabarevic

Type above and press Enter to search. Press Esc to cancel.