# Qwen3.8 Max 0902 API — pricing & specs

Qwen3.8 Max 0902, made by Alibaba, accepts text, image, and video input and returns text. On KeepRouter, Qwen3.8 Max 0902 costs $2.50 per 1M input tokens and $6.00 per 1M output tokens, billed pay-as-you-go with no monthly fee; actual request cost depends on measured token usage. Call it through a compatible KeepRouter endpoint supported by its active route, with the model id `qwen3.8-max-0902`.

| Spec | Value |
|---|---|
| Maker | Alibaba |
| Modality | Text, vision, video |
| Context window | 1,000,000 tokens |
| Max output | 131,072 tokens |
| Released | 2026-09-02 |
| Input price | $2.50 per 1M tokens |
| Output price | $6.00 per 1M tokens |
| Cached input | $0.2500 per 1M tokens |
| Capabilities | Video input, Vision (image input), Chat completions, Streaming where supported, Tool calling where supported |
| Endpoint | POST /v1/chat/completions or POST /v1/messages (compatible chat routes) |
| Model id | `qwen3.8-max-0902` |

## Call it via compatible chat endpoint

Use POST /v1/chat/completions or, on compatible chat routes, POST /v1/messages; streaming and tool support depend on the model's active route.

### cURL

```bash
curl https://keeprouter.com/v1/chat/completions \
  -H "Authorization: Bearer $KEEPROUTER_KEY" -H "Content-Type: application/json" \
  -d '{"model":"qwen3.8-max-0902","messages":[{"role":"user","content":"Hello"}]}'
```

### Python

```python
import os
from openai import OpenAI
client = OpenAI(base_url="https://keeprouter.com/v1", api_key=os.environ["KEEPROUTER_KEY"])
r = client.chat.completions.create(model="qwen3.8-max-0902", messages=[{"role":"user","content":"Hello"}])
```

### JavaScript

```js
import OpenAI from "openai";
const client = new OpenAI({ baseURL: "https://keeprouter.com/v1", apiKey: process.env.KEEPROUTER_KEY });
const r = await client.chat.completions.create({ model: "qwen3.8-max-0902", messages: [{ role: "user", content: "Hello" }] });
```

## Estimate API costs

At the current KeepRouter customer price, an example workload of 1,000 total input tokens, no cached input, and 500 output tokens per request costs approximately $0.005500 per request. At 100 requests per day, that is $16.50 over 30 days. This is a usage estimate, excluding processing fees, taxes, retries and application infrastructure. Actual usage, cache hits and supported generation durations need their own checks.

[Adjust quantities in the API cost calculator](/tools/api-cost-calculator?model=qwen3.8-max-0902).

## How to evaluate Qwen3.8 Max 0902

Qwen3.8 Max 0902 is listed on KeepRouter as text, vision, video under the exact id `qwen3.8-max-0902`. Use /v1/chat/completions for the listed route; a maker's upstream features do not automatically apply to this gateway endpoint. The published context window is 1,000,000 tokens and the published output limit is 131,072 tokens. These are limits, not a recommended request size.

**First workload:** Start with a text question plus one representative image, then repeat the same question with text only. Compare answer quality, latency and billed input/output tokens.

**Before production:** For agent use, test tool calls and structured output on this exact model route before production; an OpenAI-compatible chat endpoint alone does not prove either feature.

### Agent setup paths

### OpenCode

OpenCode documents a custom OpenAI-compatible provider. Set baseURL to https://keeprouter.com/v1 and list `qwen3.8-max-0902` as a model; validate tool behavior on the exact route. [Official setup](https://opencode.ai/docs/providers).

```json
{
  "$schema": "https://opencode.ai/config.json",
  "provider": {
    "keeprouter": {
      "npm": "@ai-sdk/openai-compatible",
      "name": "KeepRouter",
      "options": {
        "baseURL": "https://keeprouter.com/v1",
        "apiKey": "{env:KEEPROUTER_API_KEY}"
      },
      "models": {
        "qwen3.8-max-0902": {
          "name": "Qwen3.8 Max 0902"
        }
      }
    }
  }
}
```

### Continue

Continue documents provider: openai with a custom apiBase. Set apiBase to https://keeprouter.com/v1 and model to `qwen3.8-max-0902`; validate the selected feature and endpoint. [Official setup](https://docs.continue.dev/customize/model-providers/top-level/openai).

```yaml
name: KeepRouter
version: 0.0.1
schema: v1
models:
  - name: Qwen3.8 Max 0902
    provider: openai
    model: qwen3.8-max-0902
    apiBase: https://keeprouter.com/v1
    apiKey: <YOUR_KEEPROUTER_API_KEY>
```


These are documented configuration paths; model-specific tool, streaming and multimodal behavior still needs a real request test.

## Public model usage evidence

OpenRouter identifies the corresponding variant as [`qwen/qwen3.8-max-0902`](https://openrouter.ai/qwen/qwen3.8-max-0902). The figures below are OpenRouter token volume, not API call counts, KeepRouter traffic or market-wide usage.

- **Hermes Agent (agent):** 132B tokens attributed to this exact model in OpenRouter's public Apps block; observed 2026-09-28. The model page does not state the Apps block's measurement window. [OpenRouter model apps](https://openrouter.ai/qwen/qwen3.8-max-0902).
- **Portkey AI (router):** 84B tokens attributed to this exact model in OpenRouter's public Apps block; observed 2026-09-28. The model page does not state the Apps block's measurement window. [OpenRouter model apps](https://openrouter.ai/qwen/qwen3.8-max-0902).

App attribution is opt-in and reflects traffic through OpenRouter. It does not prove the app uses KeepRouter, endorse this gateway, or establish a model quality ranking. Tokenizers and windows can differ; do not add the weekly total to the app figures.

Source: OpenRouter model Apps observed 2026-09-28. The model Apps block does not state its measurement window.

## Published examples and cases

OpenRouter's [model-specific Apps view](https://openrouter.ai/qwen/qwen3.8-max-0902) publicly attributes traffic to Hermes Agent and Portkey AI. This is a public adoption signal, not a published implementation study or a KeepRouter customer claim.

## Source and verification boundary

[Official Qwen3.8 Max 0902 documentation](https://www.alibabacloud.com/help/en/model-studio/qwen3-8-max). KeepRouter's live catalog is authoritative for the customer price and enabled endpoint shown here; the maker remains authoritative for upstream model capabilities and limits.

## Pricing and implementation guides

- [Grok, Qwen, Gemma and Doubao: task-based selection](https://keeprouter.com/blog/grok-qwen-gemma-doubao-model-selection)
- [Calculate your API workload cost](https://keeprouter.com/tools/api-cost-calculator)
- [KeepRouter API keys: move from free to paid models](https://keeprouter.com/blog/keeprouter-api-key-free-to-paid)

## Guides

- [Call Qwen3.8 Max 0902 with the OpenAI SDK](https://keeprouter.com/use-cases/openai-sdk.md)

## Related models

- [Wan 2.7 Image-to-Video](https://keeprouter.com/models/wan2.7-i2v.md) by Alibaba — $0.1000 per second of video
- [Qwen 3.8 27B](https://keeprouter.com/models/qwen3.8-27b.md) by Alibaba — $0.4500 per 1M input tokens and $3.20 per 1M output tokens
- [Wan 2.7 Image-to-Video (1080p)](https://keeprouter.com/models/wan2.7-i2v-1080p.md) by Alibaba — $0.1500 per second of video
- [Qwen3.7 Max](https://keeprouter.com/models/qwen3.7-max.md) by Alibaba — $1.25 per 1M input tokens and $3.75 per 1M output tokens
- [Wan 2.7 Reference-to-Video](https://keeprouter.com/models/wan2.7-r2v.md) by Alibaba — $0.1000 per second of video
- [HappyHorse 1.0 Video Edit (1080p)](https://keeprouter.com/models/happyhorse-1.0-video-edit-1080p.md) by Alibaba — $0.2400 per second of video

## More

- [All models & pricing](https://keeprouter.com/models.md)
- [Quickstart](https://keeprouter.com/docs/quickstart.md)
- [Get an API key](https://keeprouter.com/login?returnTo=%2Fconsole%2Fkeys%3Fmodel%3Dqwen3.8-max-0902)

_Catalog facts and prices last changed 2026-10-03._
_Page content reviewed 2026-09-28._
