谷歌发布最新版Gemini音频模型,支持实时语音协作
Introducing our most advanced Gemini Audio models yet 🗣 Gemini 3.8 Live and 3.8 Live Extended Thin...
谷歌新出的Gemini音频模型,能实时处理语音和视频,比如你拍个漏水的水管,它就能给你一步步的音频指导,比之前更自然。
谷歌推出Gemini 3.8 Live和3.8 Live Extended Thin两个新模型。Gemini 3.8 Live能处理97种语言的实时对话和视觉上下文理解,例如通过摄像头指向问题区域获取维修说明。Gemini 3.8 Live Extended Thinking则能并行推理和说话,适合处理多步骤复杂任务,如规划活动。
Introducing our most advanced Gemini Audio models yet 🗣 Gemini 3.8 Live and 3.8 Live Extended Thin...
Introducing our most advanced Gemini Audio models yet 🗣 Gemini 3.8 Live and 3.8 Live Extended Thinking let you speak, collaborate, and execute tasks seamlessly, meaning conversing with AI just got a lot more natural. So, what’s the difference between these two models? Let’s break it down: — Gemini 3.8 Live is built for scale, speed, and cost efficiency. It can handle mid-sentence interruptions, transitions across 97 languages on the fly, and understands visual context. Figure out how to fix a broken bike chain, or deal with a leaky pipe just by pointing your camera at the problem area in Search Live for step-by-step audio instructions. — Gemini 3.8 Live Extended Thinking goes one step further to bring increased intelligence to your most complex tasks. It reasons and speaks in parallel, even narrating its progress as it works. This lets it handle multi-step, behind-the-scenes projects, like planning an event, without ever losing the conversational flow. Watch how Gemini 3.8 Live combines real-time video and voice inputs in Search Live to tackle hands-on DIY plumbing tasks step by step 👇 Your browser does not support the video tag. 🔗 View on Twitter 💬 20 🔄 36 ❤️ 295 👀 22644 📊 55 ⚡