Model Name
google/gemma-4-26B-A4B-itgoogle/gemma-4-26B-A4B-it
- Type: Generation
- Capabilities:
vision,reasoning
Overview
Gemma 4 26B-A4B is one of Google DeepMind’s most capable open models, built for advanced reasoning, coding, and multimodal understanding. It uses an MoE arcitechture for more efficient inference. It sits in the same general tier as Claude 4.5 Haiku and NVIDIA Nemotron 3 Super, with native function calling and structured JSON output for agentic workflows; strong image and video understanding for tasks like OCR and chart analysis; 256K context for long documents and repositories; and support for 140+ languages.
Pricing
| Priority | Input Tokens (per 1M) | Output Tokens (per 1M) |
|---|---|---|
| Realtime1 | $0.10 | $0.35 |
| Async | $0.08 | $0.25 |
| Batch (24h) | $0.05 | $0.17 |
Playground
Open this model in the Playground.
Footnotes
-
Realtime availability is limited. Doubleword is primarily a batch API. ↩