back

Gemma 4 E4B

gemma-4-e4b · Google DeepMind

MoE, 4B active params. Lightweight, reasoning and tools.

Served on NVIDIA B200 GPUs · Spain (EU)

Model source

Hugging Facegoogle/gemma-4-E4B-it

Specifications

Parameters
4B active · MoE
Hardware
MIG B200
Context
131,072 tokens
Quantization
BF16
Modalities
text
Capabilities
chat · reasoning · tools

Pricing

input

0.10

output

0.28

€ / million tokens · VAT excluded

Example (curl)

Compatible with the OpenAI SDK. Change the base URL and API key.

curl https://api.gpuflow.ai/v1/chat/completions \
  -H "Authorization: Bearer $GPUFLOW_API_KEY" \
  -d '{ "model": "gemma-4-e4b",
        "messages": [{ "role": "user", "content": "Hola" }] }'

Try it with 5 € free

Create your account, point your SDK at GPU Flow and call this model on NVIDIA B200 GPUs, in Spain.