周二,Google DeepMind 发布了 Gemini 3.6 Flash、3.5 Flash-Lite 和 3.5 Flash Cyber。Gemini 3.6 Flash 是谷歌的“主力模型”,承诺在编程、知识工作和多模态性能方面有所提升,同时将模型 token 使用量降低高达 17%,使其比前代产品 3.5 Flash 更便宜。
Gemini 3.5 Flash-Lite 是该系列中性价比最高的模型,而 3.5 Flash Cyber 则是一款专用模型,经过微调,能以合理的价格发现并修复网络安全漏洞。据谷歌称,该模型将作为有限访问试点计划的一部分,独家提供给政府和受信任的合作伙伴。
谷歌表示,这些版本的重点是为大规模构建 AI 智能体的客户提供效率、低延迟和可靠性。
此次发布引人注目,不仅在于谷歌推出了什么——针对编程、效率和网络安全进行了优化的更便宜、更快的模型——还在于它没有推出什么。此次更新并未包含人们期待已久的谷歌旗舰模型 Gemini Pro 的更新,该模型上一次更新是在二月份。
自那次发布以来,OpenAI 已经发布了 GPT-5.5 并开始推出 GPT-5.6,而 Anthropic 则推出了 Claude Opus 4.8、Claude Sonnet 5,并扩大了对前沿模型 Fable 5 的访问权限,凸显了竞争对手实验室紧张的发布节奏。
谷歌曾在五月份随 3.5 Flash 发布时预告了 Pro 版本的推出,称 Pro 版本“已在内部使用,我们期待下个月将其推出”。上周,彭博社报道称,谷歌在推出 3.5 Pro 时面临内部延迟,因为其难以达到内部性能目标。
Gemini Pro 模型通常是谷歌为复杂推理和编程任务提供的最高能力产品,而 Flash 模型则优先考虑生产应用中的更低成本和更快响应时间。
Google DeepMind 产品负责人 Logan Kilpatrick 周二表示,该公司目前正在与合作伙伴一起测试 Gemini 3.5 Pro,并希望“很快落地”。他还指出,团队已经开始了迄今为止最雄心勃勃的 Gemini 4 预训练工作。
On Tuesday, Google DeepMind released Gemini 3.6 Flash,3.5 Flash-Lite, and 3.5 Flash Cyber. Gemini 3.6 Flash is Google’s “workhorse model” that promises improved capabilities in coding, knowledge work, and multimodal performance while reducing token usage by up to 17%, making it cheaper than its predecessor 3.5 Flash.
Gemini 3.5 Flash-Lite is the most cost-effective model in the class, and 3.5 Flash Cyber is a specialized model that was fine-tuned for finding and fixing cybersecurity vulnerabilities at a decent price point. This model will be exclusively available to governments and trusted partners as part of a limited access pilot program, according to Google.
Google says the focus on these releases is to deliver efficiency, latency, and reliability to customers that are building AI agents at scale.
The launch is notable not just for what Google shipped — cheaper, faster models optimized for coding, efficiency, and cybersecurity — but for what it didn’t. The update doesn’t include the long-anticipated update to Google’s flagship model, Gemini Pro, which was last updated in February.
In the time since that launch, OpenAI has released GPT-5.5 and begun rolling out GPT-5.6, while Anthropic has launched Claude Opus 4.8, Claude Sonnet 5, and expanded access to its frontier Fable 5 model, highlighting the intense release pace of the rival labs.
Google teased the release of Pro as part of the 3.5 Flash release in May, saying the Pro version was “already being used internally, and we look forward to rolling it out next month.” Last week, Bloomberg reported that Google was facing internal delays in launching the 3.5 Pro as it struggled to meet internal performance goals.
Gemini Pro models are generally Google’s highest-capability offerings for complex reasoning and coding tasks, while Flash models prioritize lower cost and faster response times for production applications.
Google DeepMind product lead Logan Kilpatrick said Tuesday that the company is currently testing Gemini 3.5 Pro with partners and hopes to “land soon.” He also noted that the team has started its most ambitious pre-training run yet for Gemini 4.