gemini-3.1-flash-lite
Gemini 3.1 Flash-Lite is our most cost-efficient Gemini model, optimized for low latency use cases for high-volume, cost-sensitive LLM traffic.
上下文
1M
输入 / 1M
$0.25
输出 / 1M
$1.50
Modalities
协议
1- generateContent
可用性与性能
近 7 天
正在加载模型性能
能力
文本推理视觉音频长上下文缓存
计费明细
输入 Tokens
$0.25/1M
输出 Tokens
$1.50/1M
缓存输入
$0.025/1M
cached_output
$1.50/1M
音频输入
$0.5/1M
推理 Tokens
$1.50/1M
input_text
$0.25/1M
图片
$0.25/1M
output_text
$1.50/1M
input_above_threshold
$0.25/1M
cached_input_above_threshold
$0.025/1M
output_above_threshold
$1.50/1M
input_image_above_threshold
$0.25/1M
input_audio_above_threshold
$0.5/1M
价格以 USD / 1M tokens 展示(除非另注单位)。Batch API 与缓存命中会享有额外折扣,最终以工作空间实时费率为准。
