Google model product
Gemma 3 4B is a 4-billion-parameter vision-language model from Google, handling text and image input and generating text output.
Updated Aug 12, 2026. Default version: Gemma 3 4B
Technical details for the model's default version.
This profile uses the model's current scored version. Arena ratings and prices are shown separately.
Benchmark scores for Gemma 3 4B.
| VQAv2 (val) | 62.4% | 03 | 3 | 0.0% | C | |
| ECLeKTic | 4.6% | 04 | 8 | 57.1% | C | |
| Bird-SQL (dev) | 36.3% | 06 | 7 | 16.7% | C | |
| HiddenMath | 43.0% | 07 | 13 | 50.0% | C | |
| Natural2Code | 70.3% | 07 | 8 | 14.3% | C | |
| FACTS Grounding | 70.1% | 09 | 13 | 33.3% | C | |
| InfoVQA | 50.0% | 09 | 9 | 0.0% | C | |
| BIG-Bench Extra Hard | 11.0% | 10 | 11 | 10.0% | C | |
| BIG-Bench Hard | 72.2% | 11 | 21 | 50.0% | C | |
| MMMU (val) | 48.8% | 11 | 11 | 0.0% | C | |
| Global-MMLU-Lite | 54.5% | 13 | 14 | 7.7% | C | |
| IFEval | 90.2% | 15 | 65 | 78.1% | C | |
| TextVQA | 57.8% | 15 | 15 | 0.0% | C | |
| WMT24++ | 46.8% | 18 | 23 | 22.7% | C | |
| MathVista-Mini | 50.0% | 23 | 23 | 0.0% | C | |
| ChartQA | 68.8% | 24 | 24 | 0.0% | C | |
| DocVQA | 75.8% | 26 | 26 | 0.0% | C | |
| MATH | 75.6% | 27 | 71 | 62.9% | C | |
| GSM8k | 89.2% | 28 | 48 | 42.5% | C | |
| MBPP | 63.2% | 28 | 33 | 15.6% | C | |
| AI2D | 74.8% | 31 | 32 | 3.2% | C | |
| SimpleQA | 4.0% | 43 | 46 | 6.7% | C | |
| HumanEval | 71.3% | 56 | 66 | 15.4% | C | |
| LiveCodeBench | 12.6% | 72 | 73 | 1.4% | C | |
| MMLU-Pro | 43.6% | 122 | 129 | 5.5% | C | |
| GPQA | 30.8% | 221 | 234 | 5.6% | C |
Preference and agent-evaluation results for the default version.
| text | overall | 227 | 1290.8 | 4,171 | N/A | |
| text style control | overall | 257 | 1303.4 | 4,171 | N/A |
Provider-specific output speed and catalog latency for Gemma 3 4B. Runtime does not affect the capability score.
| DeepInfra | 33 tok/s | 0.2 s | 131.1K | 131.1K |
Output Speed is generated output tokens received per second. Catalog latency is reported separately from observed provider TTFT.
Official vendor API pricing appears first, followed by individual provider offers.
| OpenRouter | google/gemma-3-4b-it | global | $0.05 | $0.10 | 131.1K |
Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.
Available versions of this model. The score column identifies the version used in the overall ranking.
| Gemma 3 4B | 0.0 | 4B | 131.1K | 131.1K | No | Gemma |
Key information about Gemma 3 4B and its available data.
Gemma 3 4B is a 4-billion-parameter vision-language model from Google, handling text and image input and generating text output. It features a 128K context window, multilingual support, and open weights. Suitable for question answering, summarization, reasoning, and image understanding tasks.
Data as of 2026-08-11.
Open a comparison with the three ranked models immediately above and below this model.
Recommendations prioritize the same model type and family, then the closest LLMBoard score.
Common questions about Gemma 3 4B.
Gemma 3 4B's default version was released on Mar 12, 2025.
No official standard PAYG price is currently available for Gemma 3 4B. The lowest tracked third-party offer starts at $0.05 input and $0.10 output via OpenRouter.
Gemma 3 4B was created by Google.
The default version has a 131.1K token context window.
No. The default version is not marked as having publicly available weights.
1 provider offerings are linked to the default version.
Nearby ranked alternatives include Ministral 8B, Claude Haiku 3, Gemma 3n E4B LiteRT.