Gemini 3.5 Transcribe
Our most precise speech-to-text model yet
评分 78 ▲ 109 ◆ 1 免费增值 PH 精选
概述
Google推出的最新高精度语音转文字模型,专为实时精准转录场景设计,是Gemini系列在语音识别领域的延伸。
创新点
基于Gemini大模型底座的语音转文字能力,主打「智能感知」而非单纯识别——可能具备更强的上下文理解、标点预测、多语言混合识别等能力,区别于传统ASR模型
目标用户
开发者、企业技术团队、内容创作者、会议记录场景用户、字幕生成需求方
竞争格局
竞争格局激烈:直面OpenAI Whisper、AssemblyAI、Deepgram、Microsoft Azure Speech等成熟产品;Google凭借自有基础设施和Gemini生态有一定优势,但差异化壁垒尚待验证
标签
APISaaS语音识别AI工具人工智能开发者工具企业服务
团队
- Sundar Pichai
相关产品
GPT-6 Astra 94
OpenAI's most capable model for end-to-end work
Claude Fable 5 91
Anthropic's most capable model ever — free until June 22
AlphaGenome Atlas 88
Google's AI map of every possible human DNA mutation
WeatherNext 3 88
Our most advanced and accurate global weather AI model
Gemini Robotics 2 88
Google's AI brain for the next generation of robots
GPT-5.6 88
A new standard for intelligence and efficiency