llmboard.aiLeaderboard Center
Overall
Overall RankingOpen Models
Tools
Model DirectoryCompare Models
Capabilities
CodingReasoningMathKnowledgeInstruction Following
Price & Efficiency
Price & ValueCapability vs. PriceRuntime Performance
Modalities
Image GenerationVideo GenerationSpeech ModelsEmbeddings
Core Benchmarks
GPQAMMLU-ProAIME 2025SWE-Bench VerifiedMMLUHumanity's Last ExamLiveCodeBenchMATHHumanEvalMMMU-ProView all benchmarks
Methods
Scoring & Data
393 models668 benchmarks

Leaderboard Center

Overall RankingCodingCore BenchmarksPrice & ValueRuntime Performance

Modalities

All ModelsImage GenerationImage EditingVideo GenerationImage-to-VideoVideo EditingText-to-SpeechSpeech-to-TextEmbeddings

Data & Methods

Scoring MethodAll BenchmarksReasoningMath

Vendors

All VendorsOpenAIAnthropicGoogle
llmboard.aiCopyright 2026 llmboard.ai

Google model product

Gemma 4 E4B

5 billion effective parameters (8B with embeddings) and a 128K context window.

Updated Aug 12, 2026. Default version: Gemma 4 E4B

Compare
LLMBoard score24.4Gemma 4 E4B
Coverage80%11 benchmark families
Context window131.1KTokens
Official input priceN/AOfficial price unavailable

On this page

  • Specification
  • Capability
  • Benchmarks
  • Arena
  • Runtime
  • Pricing
  • Versions
  • About
  • Compare
  • Similar models
  • FAQ

Gemma 4 E4B Specifications

Technical details for the model's default version.

Version
Gemma 4 E4B
Released
Apr 2, 2026
Knowledge cutoff
Jan 1, 2025
Parameters
8B
Context window
131.1K
Max output
8.2K
Inputs
audio, image, text
Outputs
text
Open weights
Yes
License
Apache 2.0

Gemma 4 E4B Capability Profile

This profile uses the model's current scored version. Arena ratings and prices are shown separately.

Gemma 4 E4B category scores

Gemma 4 E4B Benchmark Results

Benchmark scores for Gemma 4 E4B.

11 rows
Columns

Show columns

BIG-Bench Extra Hard33.1%051160.0%CAug 11, 2026
MedXpertQA28.7%101218.2%CAug 11, 2026
MRCR v2 (8-needle)25.4%162125.0%CAug 11, 2026
AIME 202642.5%17185.9%CAug 11, 2026
t2-bench57.5%192318.2%CAug 11, 2026
MathVision59.5%243225.8%CAug 11, 2026
MMMLU76.6%414916.7%CAug 11, 2026
LiveCodeBench v652.0%435319.2%CAug 11, 2026
MMMU-Pro52.6%576613.8%CAug 11, 2026
MMLU-Pro69.4%8312935.9%CAug 11, 2026
GPQA58.6%16323430.5%CAug 11, 2026

Gemma 4 E4B Arena Results

Preference and agent-evaluation results for the default version.

No Arena results

The default version does not have a matching Arena result yet.

Gemma 4 E4B Runtime Performance

Provider-specific output speed and catalog latency for Gemma 4 E4B. Runtime does not affect the capability score.

No runtime data

No provider-specific speed or latency record is linked to the default version yet.

Browse runtime rankings

Output Speed is generated output tokens received per second. Catalog latency is reported separately from observed provider TTFT.

Gemma 4 E4B Pricing

Official vendor API pricing appears first, followed by individual provider offers.

Official API
N/A
Official provider
N/A
Lowest third-party
From $0.02 input, $0.10 output per 1M via Deep Infra
Tracked offerings
3
3 rows
Columns

Show columns

Deep Infragoogle/gemma-4-E4B-itglobal$0.02$0.10131.1KAug 11, 2026
NanoGPTgemma-4-e4b-itglobal$0.04$0.20131.1KAug 11, 2026
Pioneergoogle/gemma-4-E4B-itglobal$0.20$0.2032.8KAug 11, 2026

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

Gemma 4 E4B Versions

Available versions of this model. The score column identifies the version used in the overall ranking.

1 rows
Columns

Show columns

Gemma 4 E4BApr 2, 202624.48B131.1K8.2KYesApache 2.0

What is Gemma 4 E4B?

Key information about Gemma 4 E4B and its available data.

5 billion effective parameters (8B with embeddings) and a 128K context window. Supports image, text, and audio inputs. Features Per-Layer Embeddings for efficient on-device deployment while maintaining strong multimodal capabilities.

Data as of 2026-08-11.

Gemma 4 E4B vs nearby models

Open a comparison with the three ranked models immediately above and below this model.

Gemma 4 E4BvsQwen3 VL 8BGemma 4 E4BvsGPT-4.1-miniGemma 4 E4BvsLlama 3.1 405BGemma 4 E4BvsDevstral MediumGemma 4 E4BvsGPT-4oGemma 4 E4BvsQwen3 VL 4B Thinking

Models similar to Gemma 4 E4B

Recommendations prioritize the same model type and family, then the closest LLMBoard score.

#187-3.4
GO

Gemini 2.0 Flash Lite

Google

21.0 LLMBoard

DetailsCompare
#167+3.7
GO

Gemini 2.0 Flash

Google

28.1 LLMBoard

DetailsCompare
#191-3.9
GO

Gemini 1.5 Pro

Google

20.5 LLMBoard

DetailsCompare
#195-5.4
GO

Gemini 2.5 Flash Lite

Google

19.1 LLMBoard

DetailsCompare
#163+5.5
GO

Gemini 2.0 Flash Thinking

Google

29.9 LLMBoard

DetailsCompare
#203-8.7
GO

Gemma 3 27B

Google

15.8 LLMBoard

DetailsCompare

FAQ

Common questions about Gemma 4 E4B.

When was Gemma 4 E4B released?

Gemma 4 E4B's default version was released on Apr 2, 2026.

How much does Gemma 4 E4B cost?

No official standard PAYG price is currently available for Gemma 4 E4B. The lowest tracked third-party offer starts at $0.02 input and $0.10 output via Deep Infra.

Who created Gemma 4 E4B?

Gemma 4 E4B was created by Google.

What is the context window for Gemma 4 E4B?

The default version has a 131.1K token context window.

Is Gemma 4 E4B open weight?

Yes. The default version is marked as open weight under Apache 2.0.

How many API providers offer Gemma 4 E4B?

3 provider offerings are linked to the default version.

What models should I compare Gemma 4 E4B with?

Nearby ranked alternatives include Qwen3 VL 8B, GPT-4.1-mini, Llama 3.1 405B.