← model directorygoogle
Google: Gemma 4 26B A4B
google/gemma-4-26b-a4b-it
↓ runs free on your own hardware⚙ tool calling
Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...
specs & pricing
- Type
- text
- Provider
- Model ID
- google/gemma-4-26b-a4b-it
- Capabilities
- vision, tools, reasoning
- Context window
- 262K tokens
- Self-hostable
- Yes — runs on your own GPU
- Input price
- $0.046 / 1M tokens
- Output price
- $0.24 / 1M tokens
Cloud price is billed from prepaid credits when a request fails over to the cloud. Open-weight models run free on GPUs you own — the gateway routes to your nodes first.