Google 推出三款新模型,旨在提升性能、降低延迟和成本。
其中,3.6 Flash 在复杂编码任务上 token 用量最高减少 65%,3.5 Flash-Lite 速度达 350 输出 token/秒。
3.6 Flash 和 3.5 Flash-Lite 已在 Gemini 应用上线,3.5 Pro 进入合作伙伴测试。
Google has launched three new models aimed at improving performance, reducing latency, and lowering costs. Among them, 3.6 Flash reduces token usage by up...
Google has launched three new models aimed at improving performance, reducing latency, and lowering
costs. Among them, 3.6 Flash reduces token usage by up to 65% in complex encoding tasks, and 3.5 Flash-Lite reaches a speed of 350 output tokens per second. 3.6 Flash and 3.5 Flash-Lite have been launched in Gemini applications, while 3.5 Pro has entered partner testing.
Google 推出三款新模型,旨在提升性能、降低延迟和成本。
其中,3.6 Flash 在复杂编码任务上 token 用量最高减少 65%,3.5 Flash-Lite 速度达 350 输出 token/秒。
3.6 Flash 和 3.5 Flash-Lite 已在 Gemini 应用上线,3.5 Pro 进入合作伙伴测试。
公开材料称 Google 推出三款新模型,目标是提升性能并降低延迟和成本。正文明确提及 3.6 Flash、3.5 Flash-Lite 与 3.5 Pro,但与标题中的 3.5 Flash Cyber 存在名称差异。
材料显示,3.6 Flash 在复杂编码任务上的 token 用量最高减少 65%,3.5 Flash-Lite 的速度达到每秒 350 个输出 token;上述指标的测试条件与对比基准未在材料中说明。
Aioga 判断,本次更新的主要信号是 Google 同时强调模型效率、生成速度与上线进度。值得关注的是标题和正文对第三款模型的表述不一致,因此不能据此确认 3.5 Flash Cyber 的具体发布状态。
3.6 Flash 和 3.5 Flash-Lite 已在 Gemini 应用上线,意味着用户可在该应用中接触这两款模型。Aioga 判断,公开数字可能体现效率改善,但缺少测试口径,暂不宜直接推导实际成本降幅。 值得关注 Google 后续是否澄清第三款模型究竟为 3.5 Flash Cyber 还是 3.5 Pro,并补充 token 节省比例与输出速度的测试环境、对照模型及适用任务范围。
The readable text on this page was extracted from the public source and organized with attribution, publication time and the original link. Copyright remains with the original author and publisher.
Ingestion channel: Summary aggregation · Source domain: x.com
Source: X:Josh Woodward (@joshwoodward, Google Labs VP)
Original link: Open original source
Aioga archive: Open intelligence page
Content record: social-summary · Updated: 2026-07-21T15:54:29.000Z

统一接入主流 AI 模型 API,为开发、测试与生产环境提供稳定调用入口。
立即访问 api.w173.comAioga aggregates global AI updates and preserves source information for verification and citation.