加载中... --°C -- · --% · --
|
加载中... --°C -- · --% · --

Google announces Gemini 3.5 Transcribe for AI-powered speech-to-text

AI工具
Google announces Gemini 3.5 Transcribe for AI-powered speech-to-text
摘要

谷歌发布Gemini 3.5 Transcribe,一款用于语音转文字的AI模型。该模型旨在优化语音输入,可自动去除“嗯”等口头语并修正表达,输出更精炼的文本。它已支持Pixel 11的Gboard“Rambler”功能,并将推广至谷歌生态内更多产品。据谷歌介绍,其处理速度比上一代Chirp 3引擎提升约70%,实时语音错误率降至5.5%,优于Chirp 3

While we wait (possibly in vain) for Gemini 3.5 Pro to launch, Google is releasing a different model in the 3.5 branch. The company has announced Gemini 3.5 Transcribe, an AI model designed to streamline voice input by editing out "ums" and corrections, outputting polished AI text. This model already powers the Gboard "Rambler" feature on the Pixel 11, but it's about to appear throughout the Google ecosystem.

According to Google, Gemini 3.5 Transcribe is much faster and more accurate than its previous voice-to-text engine, known as Chirp 3. The new AI model should be about 70 percent faster from voice to final transcribed text, and the live-speech error rate has dropped to 5.5 percent. That's only a little better than Chirp 3, which Google measures at 7.32 percent. Still, it's a pain to fix typos when you're using voice input, so any improvement here is beneficial.

Credit: Google

Read full article

Comments

转载信息
原文: Google announces Gemini 3.5 Transcribe for AI-powered speech-to-text (2026-08-26T19:19:22)
作者: Ryan Whitwam 分类: 科技
评论 (0)
登录 后发表评论

暂无评论,来留下第一条评论吧