跳到正文
北京时间
原文
Josh Woodward· @joshwoodward · X·· 2026-07-21精选AI 评分69
AI 导读

Google 推出三款新模型,旨在提升性能、降低延迟和成本。其中,3.6 Flash 在复杂编码任务上 token 用量最高减少 65%,3.5 Flash-Lite 速度达 350 输出 token/秒。3.6 Flash 和 3.5 Flash-Lite 已在 Gemini 应用上线,3.5 Pro 进入合作伙伴测试。

推荐理由

Google一口气扔出三个Gemini Flash新模型,3.6 Flash把复杂编码token砍了65%,做Agent的成本敏感型选手可以直接收益,Cyber模型专攻漏洞修复是个有意思的细分信号。

正文

Today’s launches are all about better performance, lower latency, and a smaller bill.

  • 3.6 Flash cuts token usage by up to 65% on complex coding
  • 3.5 Flash-Lite reaches speeds of 350 output tokens/sec

Both are live in the Gemini app today!

Next up: Gemini 3.5 Pro, which has officially entered partner testing.

引用Google DeepMind@GoogleDeepMind
We’re rolling out three new models to make AI agents faster, smarter, and cheaper at scale: 🔵 Gemini 3.6 Flash: It uses fewer tokens than 3.5 Flash to deliver higher quality work at the exact same cost. 🔵 Gemini 3.5 Flash-Lite: A fast, cost-effective option for everyday tasks like processing documents and agentic search. 🔵 Gemini 3.5 Flash Cyber: A cybersecurity model built to find and patch critical software vulnerabilities.
在 X 查看被引用的帖子

来源:Josh Woodward · x.com