Google

GGemma 3n 4BOPEN

Gemma 3n 4B (Google): live API pricing, context window and the best open-source alternatives, tracked daily by olud.ai.

Context window
33K
tokens
Input price
$0.06
per M tokens
Output price
$0.12
per M tokens
Provider
Google

Prices update automatically — checked hourly against provider list prices.

See model comparisons → Compare all model prices

Benchmarks & performance

Independent benchmark scores for Gemma 3n 4B, measured by Artificial Analysis. Higher is better.

2.6×
More intelligence per dollar than Claude Opus 5 (Fast).
For the same budget, Gemma 3n 4B delivers 2.6 times more capability. Open weights also mean you can self-host it and pay nothing per token.
Intelligence index1.2
Coding index3.2
Math index14.3
GPQA29.6%
MMLU-Pro48.8%
Humanity's Last Exam4.4%
Long Context Reasoning0%
LiveCodeBench14.6%
SciCode8.1%
MATH-50077.1%
AIME13.7%
AIME 202514.3%
IFBench27.9%
τ²-Bench5%
τ-Bench Banking0.2%
Terminal-Bench0.7%
Terminal-Bench Hard2.3%
💰 Blended price$0.075 / 1M tokens
📈 Value16 intelligence points per $
Benchmark data by Artificial Analysis

About this model

Gemma 3n 4B is an open-weight AI model by Google. You can download and self-host it for free; the prices below are hosted-API list prices, tracked hourly, for when you prefer convenience over self-hosting.

Frequently asked questions

What is Gemma 3n 4B?

Gemma 3n 4B is an AI language model from Google. It is open-weight: you can download it and run it on your own hardware, for free. It scores 1.2 on the Artificial Analysis intelligence index.

Is Gemma 3n 4B free?

The weights are free and open — you can self-host Gemma 3n 4B and pay nothing per token. If you prefer a hosted API, list prices are $0.06 per million input tokens and $0.12 per million output tokens.

What is Gemma 3n 4B good at?

Independent benchmarks from Artificial Analysis give it GPQA 29.6%, MMLU-Pro 48.8%, Humanity's Last Exam 4.4%, Long Context Reasoning 0%, LiveCodeBench 14.6%, SciCode 8.1%, MATH-500 77.1%, AIME 13.7%, AIME 2025 14.3%, IFBench 27.9%, τ²-Bench 5%, τ-Bench Banking 0.2%, Terminal-Bench 0.7%, Terminal-Bench Hard 2.3%. It is particularly used for code generation.

Can I self-host Gemma 3n 4B?

Yes. Gemma 3n 4B has open weights, so you can download it and run it on your own GPU or server with tools like Ollama, vLLM or llama.cpp — with no per-token cost.

Related models

Gemini 3.1 Pro PreviewGoogleGemini 3.1 Pro Preview Custom ToolsGoogleNano Banana Pro (Gemini 3 Pro Image)GoogleGoogle Gemini Pro LatestGoogleGemini 2.5 ProGoogleGemini 2.5 Pro Preview 05-06Google