Josh Woodward· @joshwoodward · X·· 2026-07-21精选AI 评分69
AI 导读
Google 推出三款新模型,旨在提升性能、降低延迟和成本。其中,3.6 Flash 在复杂编码任务上 token 用量最高减少 65%,3.5 Flash-Lite 速度达 350 输出 token/秒。3.6 Flash 和 3.5 Flash-Lite 已在 Gemini 应用上线,3.5 Pro 进入合作伙伴测试。
推荐理由
Google一口气扔出三个Gemini Flash新模型,3.6 Flash把复杂编码token砍了65%,做Agent的成本敏感型选手可以直接收益,Cyber模型专攻漏洞修复是个有意思的细分信号。
正文
Today’s launches are all about better performance, lower latency, and a smaller bill.
- 3.6 Flash cuts token usage by up to 65% on complex coding
- 3.5 Flash-Lite reaches speeds of 350 output tokens/sec
Both are live in the Gemini app today!
Next up: Gemini 3.5 Pro, which has officially entered partner testing.
We’re rolling out three new models to make AI agents faster, smarter, and cheaper at scale: 🔵 Gemini 3.6 Flash: It uses fewer tokens than 3.5 Flash to deliver higher quality work at the exact same cost. 🔵 Gemini 3.5 Flash-Lite: A fast, cost-effective option for everyday tasks like processing documents and agentic search. 🔵 Gemini 3.5 Flash Cyber: A cybersecurity model built to find and patch critical software vulnerabilities.在 X 查看被引用的帖子
来源:Josh Woodward · x.com