Gemma 4 E2B Instruct
Open Gemma instruction model for efficient chat and self-hosted deployments
Model facts
- Context window
- 131K tokens
- Maximum output
- 16K tokens
- Cheapest paid input
- $0.02 per 1M tokens
- Cheapest paid output
- $0.1 per 1M tokens
- Open weights
- yes
- Providers
- 1
Capabilities and modalities
tool calling, structured output, attachments, temperature control, open weights, text, image, audio, video.
Artificial Analysis benchmarks
Compare this model on the full LLM benchmark leaderboard.
| Configuration | AAI | Coding | Math | Output tokens/s |
|---|---|---|---|---|
| Reasoning | 9.5 | 7.2 | not measured | not measured |
| Non-reasoning | 6.2 | not measured | not measured | not measured |
Observed price history
- 2026-08-24: $0.02 input / $0.1 output per 1M tokens
API providers
- NanoGPT — model id gemma-4-e2b-it; $0.02 input / $0.1 output per 1M tokens; provider documentation