gemini-3.1-flash-lite

Gemini 3.1 Flash-Lite is our most cost-efficient Gemini model, optimized for low latency use cases for high-volume, cost-sensitive LLM traffic.

LiveGoogle1 protocol1M contextStream cancellation unsupported
Ngữ cảnh
1M
Input / 1M
$0.25
Output / 1M
$1.50
Loại dữ liệu
Protocol
1
  • generateContent

Status & performance

Last 7 days
Loading model performance

Capabilities

Văn bảnSuy luậnThị giácÂm thanhNgữ cảnh dàiBộ nhớ đệm

Bảng giá

Token đầu vào
$0.25/1M
Token đầu ra
$1.50/1M
Cached input
$0.025/1M
cached_output
$1.50/1M
Audio input
$0.5/1M
Reasoning tokens
$1.50/1M
input_text
$0.25/1M
Hình ảnh
$0.25/1M
output_text
$1.50/1M
input_above_threshold
$0.25/1M
cached_input_above_threshold
$0.025/1M
output_above_threshold
$1.50/1M
input_image_above_threshold
$0.25/1M
input_audio_above_threshold
$0.5/1M

Prices in USD per 1M tokens unless noted otherwise. Batch calls and cache hits receive additional discounts; live rates apply in the Workspace.