Skip to content

Gemini 3.1 Flash Lite

ModelActive
by Google · family “gemini-flash-lite”

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic... (description from the OpenRouter listing)

in: textin: imagein: videoin: filein: audioout: textReasoningTool useStructured output
Epoch Capabilities Index
144.4 #83 of 268
90% CI 142.3 – 145.9
Listing ↗
Released
3 Mar 2026
Context window
1.05M tokens
Max output
66K tokens
Input $ / 1M tokens
$0.25
Output $ / 1M tokens
$1.50
Cached input $ / 1M
$0.025
Each benchmark shown separately with its own source

Benchmark results (9)

BenchmarkDomainScorevs best recordedSettingRunSource
LMCAagents35.0%
51%
high—External ↗
DTBenchreasoning76.8%
78%
high—External ↗
FrontierMath-Tiers-1-3-v2-Privatemath27.7% ±2.7
30%
high30 Aug 2026Eval log ↗
Chess Puzzlesgames25.0% ±4.4
35%
low6 Aug 2026Eval log ↗
DeepResearch Benchagents37.3%
67%
low—External ↗
HLEknowledge8.6%
16%
——External ↗
WeirdMLcoding52.2% ±0.0
56%
——External ↗
OTIS Mock AIME 2024-2025math80.0% ±6.0
80%
high6 Aug 2026Eval log ↗
GPQA diamondscience81.8% ±2.7
85%
high6 Aug 2026Eval log ↗

Source: Epoch AI Benchmarking Hub (CC BY 4.0). “External” rows are leaderboard results Epoch collects from third parties. Best reported setting per benchmark is shown.

API list price over time

Price history

$ per 1M tokens

5 recorded prices on 6 Jun 2026 (OpenRouter listing + Internet Archive snapshots); re-read on every data refresh, most recently 6h ago. Steps show when the price changed.

Events