Skip to content
ModelPriceLab

Search model prices

Find source-linked rates by model, vendor, or platform.

Model Services/Google AI/Gemma 4 26B A4B IT
google

Gemma 4 26B A4B IT

Active

Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...

View current API pricing for this model
Input, output, and cached rates with source and verification time, plus same-workload cost examples.
Go to pricing page →
Input / 1M tokens
$0.07
Output / 1M tokens
$0.23
Record updated 2026-10-11
Modality:LLM
Context:262,144 tokens
Release:Apr 2, 2026

Pricing

Platform
Scope
Input
Output
Details
OpenRouter
Unified API Router
Standard
$0.07
$0.23
262k context · 236k max output
Official
Standard
-
-
262k context · 33k max output

How much are you overpaying?

Enter your monthly spend on this model to see what the same workload could cost on cheaper alternatives.

$
Enter a monthly spend to see the listed-rate comparison.

Latest changes

No recent changes.

Benchmarks

No benchmark data available.

⚡Gemma 4 26B A4B IT Comparisons & Scenario Guides

❓Gemma 4 26B A4B IT Frequently Asked Questions

How much does Gemma 4 26B A4B IT cost per 1M tokens?▼

Gemma 4 26B A4B IT is priced at $0.07 per 1M input tokens and $0.23 per 1M output tokens, with cached input at $0.04/1M on openrouter. Check provider terms before deployment.

What is the context window for Gemma 4 26B A4B IT?▼

Gemma 4 26B A4B IT supports up to 262,144 tokens context with a max completion limit of 235,929 tokens.

What are cost-saving alternatives to Gemma 4 26B A4B IT?▼

ModelPriceLab automatically calculates cheaper alternatives in the same model class, showing monthly spend differences and migration guides.

AI Cost Optimization

Does this change your best option?

Check your models and usage for free. See the relevant official changes, estimated cost, and a practical next step.