Google 推出三款新模型,旨在提升性能、降低延迟和成本。其中,3.6 Flash 在复杂编码任务上 token 用量最高减少 65%,3.5 Flash-Lite 速度达 350 输出 token/秒。3.6 Flash 和 3.5 Flash-Lite 已在 Gemini 应用上线,3.5 Pro 进入合作伙伴测试。
模型
·X:Josh Woodward (@joshwoodward, Google Labs VP)
Google 发布三款新模型:3.6 Flash、3.5 Flash-Lite 与 3.5 Flash Cyber
— Today's launches are all about better performance, lower latency, and a smaller bill. + 3.6 Flash c…