# Gemma 4 26B A4B IT API — pricing & specs

Gemma 4 26B A4B IT is Google's instruction-tuned mixture-of-experts text and image model with open weights. On KeepRouter, Gemma 4 26B A4B IT costs $0.1000 per 1M input tokens and $0.3000 per 1M output tokens, billed pay-as-you-go with no monthly fee; actual request cost depends on measured token usage. Call it through a compatible KeepRouter endpoint supported by its active route, with the model id `gemma-4-26b-a4b-it`.

| Spec | Value |
|---|---|
| Maker | Google DeepMind |
| Modality | Text, vision |
| Open weights | Yes |
| Input price | $0.1000 per 1M tokens |
| Output price | $0.3000 per 1M tokens |
| Capabilities | Vision (image input), Chat completions, Streaming where supported, Tool calling where supported |
| Endpoint | POST /v1/chat/completions |
| Model id | `gemma-4-26b-a4b-it` |

## Call it via /v1/chat/completions

Use POST /v1/chat/completions or, on compatible chat routes, POST /v1/messages; streaming and tool support depend on the model's active route.

### cURL

```bash
curl https://keeprouter.com/v1/chat/completions \
  -H "Authorization: Bearer $KEEPROUTER_KEY" -H "Content-Type: application/json" \
  -d '{"model":"gemma-4-26b-a4b-it","messages":[{"role":"user","content":"Hello"}]}'
```

### Python

```python
import os
from openai import OpenAI
client = OpenAI(base_url="https://keeprouter.com/v1", api_key=os.environ["KEEPROUTER_KEY"])
r = client.chat.completions.create(model="gemma-4-26b-a4b-it", messages=[{"role":"user","content":"Hello"}])
```

### JavaScript

```js
import OpenAI from "openai";
const client = new OpenAI({ baseURL: "https://keeprouter.com/v1", apiKey: process.env.KEEPROUTER_KEY });
const r = await client.chat.completions.create({ model: "gemma-4-26b-a4b-it", messages: [{ role: "user", content: "Hello" }] });
```

## Estimate API costs

At the current KeepRouter customer price, an example workload of 1,000 total input tokens, no cached input, and 500 output tokens per request costs approximately $0.000250 per request. At 100 requests per day, that is $0.7500 over 30 days. This is a usage estimate, excluding processing fees, taxes, retries and application infrastructure. Actual usage, cache hits and supported generation durations need their own checks.

[Adjust quantities in the API cost calculator](/tools/api-cost-calculator?model=gemma-4-26b-a4b-it).

## Model identity and official sources

Sources checked 2026-10-03.

### Bounded generation checked

On October 3, 2026, this exact ID returned HTTP 200, visible text and final token usage through /v1/chat/completions. These small checks do not establish task accuracy, full-context capacity, media compatibility or sustained throughput.

[KeepRouter · API reference](https://keeprouter.com/api/docs)

### Model and evaluation task

Gemma 4 26B A4B is a mixture-of-experts model with about 25.2B total and 3.8B active parameters. Google documents a 256K model context; deployment limits may be lower. This page deliberately omits an unverified serving-context guarantee.

[Google DeepMind · model documentation](https://ai.google.dev/gemma/docs/core/model_card_4)

### API contract and migration

Use Chat Completions with this exact public model ID. Maker-native tools, audio generation and media encodings are not automatically exposed by a compatibility route. Start with text, then test each required media and tool workflow separately.

[Google DeepMind · model documentation](https://ai.google.dev/gemma/docs/core/model_card_4)

### Customer price and verification scope

Use this page's live KeepRouter USD input, output and, when shown, cached-input prices. Cached tokens are part of prompt usage and must not be counted twice. Compare spend including reasoning and retries. Reviewed documentation and catalog presence do not certify full-context capacity, media handling or throughput.

[KeepRouter · pricing and usage](https://keeprouter.com/models) · [KeepRouter · workload calculator](https://keeprouter.com/tools/api-cost-calculator)

## How to evaluate Gemma 4 26B A4B IT

Gemma 4 26B A4B IT is listed on KeepRouter as text, vision under the exact id `gemma-4-26b-a4b-it`. Use /v1/chat/completions for the listed route; a maker's upstream features do not automatically apply to this gateway endpoint. No context limit is asserted here because the exact upstream limit has not been verified for this catalog entry.

**First workload:** Start with a text question plus one representative image, then repeat the same question with text only. Compare answer quality, latency and billed input/output tokens.

**Before production:** For agent use, test tool calls and structured output on this exact model route before production; an OpenAI-compatible chat endpoint alone does not prove either feature.

### Agent setup paths

### OpenCode

OpenCode documents a custom OpenAI-compatible provider. Set baseURL to https://keeprouter.com/v1 and list `gemma-4-26b-a4b-it` as a model; validate tool behavior on the exact route. [Official setup](https://opencode.ai/docs/providers).

```json
{
  "$schema": "https://opencode.ai/config.json",
  "provider": {
    "keeprouter": {
      "npm": "@ai-sdk/openai-compatible",
      "name": "KeepRouter",
      "options": {
        "baseURL": "https://keeprouter.com/v1",
        "apiKey": "{env:KEEPROUTER_API_KEY}"
      },
      "models": {
        "gemma-4-26b-a4b-it": {
          "name": "Gemma 4 26B A4B IT"
        }
      }
    }
  }
}
```

### Continue

Continue documents provider: openai with a custom apiBase. Set apiBase to https://keeprouter.com/v1 and model to `gemma-4-26b-a4b-it`; validate the selected feature and endpoint. [Official setup](https://docs.continue.dev/customize/model-providers/top-level/openai).

```yaml
name: KeepRouter
version: 0.0.1
schema: v1
models:
  - name: Gemma 4 26B A4B IT
    provider: openai
    model: gemma-4-26b-a4b-it
    apiBase: https://keeprouter.com/v1
    apiKey: <YOUR_KEEPROUTER_API_KEY>
```


These are documented configuration paths; model-specific tool, streaming and multimodal behavior still needs a real request test.

## Public model usage evidence

No public, exact-variant usage figure has been verified for this KeepRouter model id. Missing data is not zero usage; family-level or maker-wide traffic is not presented as this model's traffic.

## Published examples and cases

No exact-model customer case has been verified for this entry. The workload above is an evaluation recipe, not a claim of a public deployment.

## Source and verification boundary

[Official Gemma 4 26B A4B IT documentation](https://ai.google.dev/gemma/docs/core/model_card_4). KeepRouter's live catalog is authoritative for the customer price and enabled endpoint shown here; the maker remains authoritative for upstream model capabilities and limits.

## Pricing and implementation guides

- [Grok, Qwen, Gemma and Doubao: task-based selection](https://keeprouter.com/blog/grok-qwen-gemma-doubao-model-selection)
- [Calculate your API workload cost](https://keeprouter.com/tools/api-cost-calculator)
- [KeepRouter API keys: move from free to paid models](https://keeprouter.com/blog/keeprouter-api-key-free-to-paid)

## Guides

- [Call Gemma 4 26B A4B IT with the OpenAI SDK](https://keeprouter.com/use-cases/openai-sdk.md)

## Related models

- [Gemma 4 31B](https://keeprouter.com/models/gemma-4-31B-it.md) by Google DeepMind — $0.1200 per 1M input tokens and $0.3500 per 1M output tokens
- [gemini-embedding-001](https://keeprouter.com/models/gemini-embedding-001.md) by Google DeepMind — $0.1500 per 1M input tokens and $0.6000 per 1M output tokens
- [Veo 3.1](https://keeprouter.com/models/veo-3.1.md) by Google DeepMind — $0.4000 per second of video
- [Gemini 3.8 Flash](https://keeprouter.com/models/gemini-3.8-flash.md) by Google DeepMind — $0.7500 per 1M input tokens and $3.75 per 1M output tokens
- [Veo 3.1 (1080p)](https://keeprouter.com/models/veo-3.1-1080p.md) by Google DeepMind — $0.4000 per second of video
- [Gemini 3.7 Flash](https://keeprouter.com/models/gemini-3.7-flash.md) by Google DeepMind — $0.7500 per 1M input tokens and $3.75 per 1M output tokens

## More

- [All models & pricing](https://keeprouter.com/models.md)
- [Quickstart](https://keeprouter.com/docs/quickstart.md)
- [Get an API key](https://keeprouter.com/login?returnTo=%2Fconsole%2Fkeys%3Fmodel%3Dgemma-4-26b-a4b-it)

_Catalog facts and prices last changed 2026-10-03._
_Page content reviewed 2026-10-03._
