Josh Woodward · @joshwoodward · X·2026-07-21 23:54·45天前
AI 导读

Google 推出三款新模型,旨在提升性能、降低延迟和成本。其中,3.6 Flash 在复杂编码任务上 token 用量最高减少 65%,3.5 Flash-Lite 速度达 350 输出 token/秒。3.6 Flash 和 3.5 Flash-Lite 已在 Gemini 应用上线,3.5 Pro 进入合作伙伴测试。

Josh Woodward@joshwoodward
同事件
69AI 编辑部评分,满分 100
2026-07-21 23:54· 45天前
AI 导读

Google 推出三款新模型,旨在提升性能、降低延迟和成本。其中,3.6 Flash 在复杂编码任务上 token 用量最高减少 65%,3.5 Flash-Lite 速度达 350 输出 token/秒。3.6 Flash 和 3.5 Flash-Lite 已在 Gemini 应用上线,3.5 Pro 进入合作伙伴测试。

Today’s launches are all about better performance, lower latency, and a smaller bill.

• 3.6 Flash cuts token usage by up to 65% on complex coding • 3.5 Flash-Lite reaches speeds of 350 output tokens/sec

Both are live in the Gemini app today!

Next up: Gemini 3.5 Pro, which has officially entered partner testing.

Google DeepMindWe’re rolling out three new models to make AI agents faster, smarter, and cheaper at scale: 🔵 Gemini 3.6 Flash: It uses fewer tokens than 3.5 Flash to deliver ...