Gemma 4 26B A4B is built for developers who need scalable performance without sacrificing core capabilities.Crucially, it retains the massive 256K-token context window of the 31B model, making it highly competitive for long-context RAG and processing extensive, image-rich document datasets.
function_callinghttps://api.novita.ai/v3/openai/v1/chat/completions| Provider | Input | Output | Since launch |
|---|---|---|---|
| DeepInfracheapest | $0.07 | $0.34 | — |
| OpenRouter | $0.09 | $0.3 | ↓31% |
| Novita AI | $0.13 | $0.4 | — |
| Vercel AI Gateway | $0.13 | $0.4 | — |
| Model | Context | Max out | Pricing |
|---|---|---|---|
| Gemma 4 26B A4B | 262K | 131K | $0.13$0.4 |
| Gemma 4 31B | 262K | 131K | $0.14$0.4 |
/v1/models/novita/google/gemma-4-26b-a4b-it