Arena Mirror (China)
Google

Gemini 2.5 Flash Lite Preview 09.2025 No Thinking

Google2025-07-22

Gemini 2.5 Flash-Lite 是 Gemini 2.5 系列中的轻量级推理模型,针对超低延迟和成本效率优化,提供更高吞吐量和更快的 token 生成速度。

模态
上下文
1.05Mtokens
最大输出 66K tokens
能力
工具调用 (Tools)
结构化输出
并行工具调用
视觉输入
音频输入
多服务商价格对比数据来自 OpenRouter
5 个服务商1 USD ≈ 6.7569 CNY · 更新于 2026-08-19
服务商量化精度上下文输入/M输出/M缓存读取/M近1天可用率
Google AI StudioGoogle AI Studio (Flex)1.05M¥0.34¥1.35¥0.0399.4%
GoogleGoogle Vertex (EU)1.05M¥0.68¥2.70¥0.0799.9%
GoogleGoogle Vertex1.05M¥0.68¥2.70¥0.0799.7%
Google AI StudioGoogle AI Studio1.05M¥0.68¥2.70¥0.0799.4%
Google AI StudioGoogle AI Studio (Priority)1.05M¥1.22¥4.86¥0.1299.4%
Google AI StudioGoogle AI Studio (Flex)
¥0.34
输入/M
¥1.35
输出/M
¥0.03
缓存/M
99.4%
可用率
GoogleGoogle Vertex (EU)
¥0.68
输入/M
¥2.70
输出/M
¥0.07
缓存/M
99.9%
可用率
GoogleGoogle Vertex
¥0.68
输入/M
¥2.70
输出/M
¥0.07
缓存/M
99.7%
可用率
Google AI StudioGoogle AI Studio
¥0.68
输入/M
¥2.70
输出/M
¥0.07
缓存/M
99.4%
可用率
Google AI StudioGoogle AI Studio (Priority)
¥1.22
输入/M
¥4.86
输出/M
¥0.12
缓存/M
99.4%
可用率