Gemma 4 26B A4B IT API — pricing & specs
Gemma 4 26B A4B IT is Google's instruction-tuned mixture-of-experts text and image model with open weights. On KeepRouter, Gemma 4 26B A4B IT costs $0.1000 per 1M input tokens and $0.3000 per 1M output tokens, billed pay-as-you-go with no monthly fee. Call it through a compatible KeepRouter endpoint supported by its active route, with the model id gemma-4-26b-a4b-it.
| Maker | Google DeepMind |
|---|---|
| Modality | Text, vision |
| Open weights | Yes |
| Input price | $0.1000 per 1M tokens |
| Output price | $0.3000 per 1M tokens |
| Capabilities | Vision (image input), Chat completions, Streaming where supported, Tool calling where supported |
| Endpoint | POST /v1/chat/completions |
| Model id | gemma-4-26b-a4b-it |
How pricing works for Gemma 4 26B A4B IT
Gemma 4 26B A4B IT is billed per token — $0.1000 per 1M input tokens and $0.3000 per 1M output tokens. The published price is pay-as-you-go, with no monthly fee; actual request cost depends on measured token usage.
Calling Gemma 4 26B A4B IT on KeepRouter
Point a compatible client at the supported KeepRouter endpoint and set the model to gemma-4-26b-a4b-it. KeepRouter preserves the client-facing request shape while handling upstream routing or translation. Use POST /v1/chat/completions or, on compatible chat routes, POST /v1/messages; streaming and tool support depend on the model's active route.
cURL
curl https://keeprouter.com/v1/chat/completions \
-H "Authorization: Bearer $KEEPROUTER_KEY" -H "Content-Type: application/json" \
-d '{"model":"gemma-4-26b-a4b-it","messages":[{"role":"user","content":"Hello"}]}'Python
import os
from openai import OpenAI
client = OpenAI(base_url="https://keeprouter.com/v1", api_key=os.environ["KEEPROUTER_KEY"])
r = client.chat.completions.create(model="gemma-4-26b-a4b-it", messages=[{"role":"user","content":"Hello"}])JavaScript
import OpenAI from "openai";
const client = new OpenAI({ baseURL: "https://keeprouter.com/v1", apiKey: process.env.KEEPROUTER_KEY });
const r = await client.chat.completions.create({ model: "gemma-4-26b-a4b-it", messages: [{ role: "user", content: "Hello" }] });Estimate API costs
At the current KeepRouter customer price, an example workload of 1,000 total input tokens, no cached input, and 500 output tokens per request costs approximately $0.000250 per request. At 100 requests per day, that is $0.7500 over 30 days. This is a usage estimate, excluding processing fees, taxes, retries and application infrastructure. Actual usage, cache hits and supported generation durations need their own checks.
Adjust quantities in the API cost calculator.
Model identity and official sources
Sources checked 2026-10-03.
Bounded generation checked
On October 3, 2026, this exact ID returned HTTP 200, visible text and final token usage through /v1/chat/completions. These small checks do not establish task accuracy, full-context capacity, media compatibility or sustained throughput.
Model and evaluation task
Gemma 4 26B A4B is a mixture-of-experts model with about 25.2B total and 3.8B active parameters. Google documents a 256K model context; deployment limits may be lower. This page deliberately omits an unverified serving-context guarantee.
Google DeepMind · model documentation
API contract and migration
Use Chat Completions with this exact public model ID. Maker-native tools, audio generation and media encodings are not automatically exposed by a compatibility route. Start with text, then test each required media and tool workflow separately.
Google DeepMind · model documentation
Customer price and verification scope
Use this page's live KeepRouter USD input, output and, when shown, cached-input prices. Cached tokens are part of prompt usage and must not be counted twice. Compare spend including reasoning and retries. Reviewed documentation and catalog presence do not certify full-context capacity, media handling or throughput.
KeepRouter · pricing and usage · KeepRouter · workload calculator
How to evaluate Gemma 4 26B A4B IT
Gemma 4 26B A4B IT is listed on KeepRouter as text, vision under the exact id gemma-4-26b-a4b-it. Use /v1/chat/completions for the listed route; a maker's upstream features do not automatically apply to this gateway endpoint. No context limit is asserted here because the exact upstream limit has not been verified for this catalog entry.
First workload: Start with a text question plus one representative image, then repeat the same question with text only. Compare answer quality, latency and billed input/output tokens.
Before production: For agent use, test tool calls and structured output on this exact model route before production; an OpenAI-compatible chat endpoint alone does not prove either feature.
Agent setup paths
OpenCode
OpenCode documents a custom OpenAI-compatible provider. Set baseURL to https://keeprouter.com/v1 and list gemma-4-26b-a4b-it as a model; validate tool behavior on the exact route. Official setup.
{
"$schema": "https://opencode.ai/config.json",
"provider": {
"keeprouter": {
"npm": "@ai-sdk/openai-compatible",
"name": "KeepRouter",
"options": {
"baseURL": "https://keeprouter.com/v1",
"apiKey": "{env:KEEPROUTER_API_KEY}"
},
"models": {
"gemma-4-26b-a4b-it": {
"name": "Gemma 4 26B A4B IT"
}
}
}
}
}Continue
Continue documents provider: openai with a custom apiBase. Set apiBase to https://keeprouter.com/v1 and model to gemma-4-26b-a4b-it; validate the selected feature and endpoint. Official setup.
name: KeepRouter
version: 0.0.1
schema: v1
models:
- name: Gemma 4 26B A4B IT
provider: openai
model: gemma-4-26b-a4b-it
apiBase: https://keeprouter.com/v1
apiKey: <YOUR_KEEPROUTER_API_KEY>These are documented configuration paths; model-specific tool, streaming and multimodal behavior still needs a real request test.
Public model usage evidence
No public, exact-variant usage figure has been verified for this KeepRouter model id. Missing data is not zero usage; family-level or maker-wide traffic is not presented as this model's traffic.
Published examples and cases
No exact-model customer case has been verified for this entry. The workload above is an evaluation recipe, not a claim of a public deployment.
Source and verification boundary
Official Gemma 4 26B A4B IT documentation. KeepRouter's live catalog is authoritative for the customer price and enabled endpoint shown here; the maker remains authoritative for upstream model capabilities and limits.
Pricing and implementation guides
- Grok, Qwen, Gemma and Doubao: task-based selection
- Calculate your API workload cost
- KeepRouter API keys: move from free to paid models
Guides
Related models
- Gemma 4 31B by Google DeepMind — $0.1200 per 1M input tokens and $0.3500 per 1M output tokens
- gemini-embedding-001 by Google DeepMind — $0.1500 per 1M input tokens and $0.6000 per 1M output tokens
- Veo 3.1 by Google DeepMind — $0.4000 per second of video
- Gemini 3.8 Flash by Google DeepMind — $0.7500 per 1M input tokens and $3.75 per 1M output tokens
- Veo 3.1 (1080p) by Google DeepMind — $0.4000 per second of video
- Gemini 3.7 Flash by Google DeepMind — $0.7500 per 1M input tokens and $3.75 per 1M output tokens
All models & pricing · Quickstart · KeepRouter vs OpenRouter · Glossary · Get an API key
Catalog facts and prices last changed .
Page content reviewed .