Gemma 4 26B A4B

26B / 4B

Gemma 4 26B A4B is built for developers who need scalable performance without sacrificing core capabilities.Crucially, it retains the massive 256K-token context window of the 31B model, making it highly competitive for long-context RAG and processing extensive, image-rich document datasets.

Context
262K
Max output
131K
Input price
$0.13/1M tokens
Output price
$0.4/1M tokens

Capabilities

vision
tool call
structured output
reasoning
json mode
streaming
fine tuning
batch

Details

Model IDgoogle/gemma-4-26b-a4b-it
Provider Novita AI
Creator Google AI
Familygemma-4
License
Parameters26B / 4B
Statusactive
Input modalitiestext, image
Output modalitiestext
Architecture
Knowledge cutoff
Training data cutoff
Release date
Deprecation date
Typechat
Reasoning tokens
Max input
Open weight
Sourceofficial
Last updated

Tools

Function Callingfunction_calling
Call external functions and APIs

Endpoints

Chat CompletionsPOST
Generate chat responses with messageshttps://api.novita.ai/v3/openai/v1/chat/completions

Pricing

Input
$0.13
Output
$0.4
Cache write
Cache read
Batch in
Batch out

Price across providers

ProviderInputOutputSince launch
DeepInfracheapest$0.07$0.34
OpenRouter$0.09$0.3↓31%
Novita AI$0.13$0.4
Vercel AI Gateway$0.13$0.4

Family Comparison: gemma-4

ModelContextMax outPricing
Gemma 4 26B A4B262K131K$0.13$0.4
Gemma 4 31B262K131K$0.14$0.4

API

GET/v1/models/novita/google/gemma-4-26b-a4b-it