Google

GGemini 3.1 Flash LiteAPI

Gemini 3.1 Flash Lite (Google): live API pricing, context window and the best open-source alternatives, tracked daily by olud.ai.

Context window
1M
tokens
Input price
$0.25
per M tokens
Output price
$1.5
per M tokens
Provider
Google

Prices update automatically — checked hourly against provider list prices.

See open-source alternatives → Compare all model prices

Benchmarks & performance

Independent benchmark scores for Gemini 3.1 Flash Lite, measured by Artificial Analysis. Higher is better.

Intelligence index25
Coding index34.7
GPQA82.2%
Humanity's Last Exam16.2%
Long Context Reasoning65.3%
SciCode41.9%
IFBench77.2%
τ²-Bench31.3%
τ-Bench Banking8.7%
Terminal-Bench31.1%
Terminal-Bench Hard24.2%
⚡ Speed309.5 tokens/sec
⏱ Latency4.92s to first token
💰 Blended price$0.563 / 1M tokens
📈 Value44.4 intelligence points per $
Benchmark data by Artificial Analysis

About this model

Gemini 3.1 Flash Lite is a commercial AI model by Google. The specifications below are tracked automatically: pricing is refreshed hourly from public list prices, so the numbers on this page reflect the current cost of using the model through its API.

Frequently asked questions

What is Gemini 3.1 Flash Lite?

Gemini 3.1 Flash Lite is an AI language model from Google. It is a proprietary model, available through an API. It scores 25 on the Artificial Analysis intelligence index.

Is Gemini 3.1 Flash Lite free?

Gemini 3.1 Flash Lite is not free: it costs $0.25 per million input tokens and $1.5 per million output tokens. Open-weight alternatives can be self-hosted at no per-token cost.

What is Gemini 3.1 Flash Lite good at?

Independent benchmarks from Artificial Analysis give it GPQA 82.2%, Humanity's Last Exam 16.2%, Long Context Reasoning 65.3%, SciCode 41.9%, IFBench 77.2%, τ²-Bench 31.3%, τ-Bench Banking 8.7%, Terminal-Bench 31.1%, Terminal-Bench Hard 24.2%. It is particularly used for code generation.

How fast is Gemini 3.1 Flash Lite?

It generates about 309.5 tokens per second, with a median 4.92s delay before the first token. Measured independently by Artificial Analysis.

Related models

Gemini 3.1 Pro PreviewGoogleGemini 3.1 Pro Preview Custom ToolsGoogleNano Banana Pro (Gemini 3 Pro Image)GoogleGoogle Gemini Pro LatestGoogleGemini 2.5 ProGoogleGemini 2.5 Pro Preview 05-06Google