Qwen3.8 Max 0902 API — pricing & specs

Qwen3.8 Max 0902, made by Alibaba, accepts text, image, and video input and returns text. On KeepRouter, Qwen3.8 Max 0902 costs $2.50 per 1M input tokens and $6.00 per 1M output tokens, billed pay-as-you-go with no monthly fee. Call it through a compatible KeepRouter endpoint supported by its active route, with the model id qwen3.8-max-0902.

MakerAlibaba
ModalityText, vision, video
Context window1,000,000 tokens
Max output131,072 tokens
Released2026-09-02
Input price$2.50 per 1M tokens
Output price$6.00 per 1M tokens
Cached input$0.2500 per 1M tokens
CapabilitiesVideo input, Vision (image input), Chat completions, Streaming where supported, Tool calling where supported
EndpointPOST /v1/chat/completions or POST /v1/messages (compatible chat routes)
Model idqwen3.8-max-0902

How pricing works for Qwen3.8 Max 0902

Qwen3.8 Max 0902 is billed per token — $2.50 per 1M input tokens and $6.00 per 1M output tokens, with cached input at $0.2500 per 1M tokens. The published price is pay-as-you-go, with no monthly fee; actual request cost depends on measured token usage.

Calling Qwen3.8 Max 0902 on KeepRouter

Point a compatible client at the supported KeepRouter endpoint and set the model to qwen3.8-max-0902. KeepRouter preserves the client-facing request shape while handling upstream routing or translation. Use POST /v1/chat/completions or, on compatible chat routes, POST /v1/messages; streaming and tool support depend on the model's active route.

cURL

curl https://keeprouter.com/v1/chat/completions \
  -H "Authorization: Bearer $KEEPROUTER_KEY" -H "Content-Type: application/json" \
  -d '{"model":"qwen3.8-max-0902","messages":[{"role":"user","content":"Hello"}]}'

Python

import os
from openai import OpenAI
client = OpenAI(base_url="https://keeprouter.com/v1", api_key=os.environ["KEEPROUTER_KEY"])
r = client.chat.completions.create(model="qwen3.8-max-0902", messages=[{"role":"user","content":"Hello"}])

JavaScript

import OpenAI from "openai";
const client = new OpenAI({ baseURL: "https://keeprouter.com/v1", apiKey: process.env.KEEPROUTER_KEY });
const r = await client.chat.completions.create({ model: "qwen3.8-max-0902", messages: [{ role: "user", content: "Hello" }] });

Estimate API costs

At the current KeepRouter customer price, an example workload of 1,000 total input tokens, no cached input, and 500 output tokens per request costs approximately $0.005500 per request. At 100 requests per day, that is $16.50 over 30 days. This is a usage estimate, excluding processing fees, taxes, retries and application infrastructure. Actual usage, cache hits and supported generation durations need their own checks.

Adjust quantities in the API cost calculator.

How to evaluate Qwen3.8 Max 0902

Qwen3.8 Max 0902 is listed on KeepRouter as text, vision, video under the exact id qwen3.8-max-0902. Use /v1/chat/completions for the listed route; a maker's upstream features do not automatically apply to this gateway endpoint. The published context window is 1,000,000 tokens and the published output limit is 131,072 tokens. These are limits, not a recommended request size.

First workload: Start with a text question plus one representative image, then repeat the same question with text only. Compare answer quality, latency and billed input/output tokens.

Before production: For agent use, test tool calls and structured output on this exact model route before production; an OpenAI-compatible chat endpoint alone does not prove either feature.

Agent setup paths

OpenCode

OpenCode documents a custom OpenAI-compatible provider. Set baseURL to https://keeprouter.com/v1 and list qwen3.8-max-0902 as a model; validate tool behavior on the exact route. Official setup.

{
  "$schema": "https://opencode.ai/config.json",
  "provider": {
    "keeprouter": {
      "npm": "@ai-sdk/openai-compatible",
      "name": "KeepRouter",
      "options": {
        "baseURL": "https://keeprouter.com/v1",
        "apiKey": "{env:KEEPROUTER_API_KEY}"
      },
      "models": {
        "qwen3.8-max-0902": {
          "name": "Qwen3.8 Max 0902"
        }
      }
    }
  }
}

Continue

Continue documents provider: openai with a custom apiBase. Set apiBase to https://keeprouter.com/v1 and model to qwen3.8-max-0902; validate the selected feature and endpoint. Official setup.

name: KeepRouter
version: 0.0.1
schema: v1
models:
  - name: Qwen3.8 Max 0902
    provider: openai
    model: qwen3.8-max-0902
    apiBase: https://keeprouter.com/v1
    apiKey: <YOUR_KEEPROUTER_API_KEY>

These are documented configuration paths; model-specific tool, streaming and multimodal behavior still needs a real request test.

Public model usage evidence

OpenRouter identifies the corresponding variant as qwen/qwen3.8-max-0902. The figures below are OpenRouter token volume, not API call counts, KeepRouter traffic or market-wide usage.

App attribution is opt-in and reflects traffic through OpenRouter. It does not prove the app uses KeepRouter, endorse this gateway, or establish a model quality ranking. Tokenizers and windows can differ; do not add the weekly total to the app figures.

Source: OpenRouter model Apps observed 2026-09-28. The model Apps block does not state its measurement window.

Published examples and cases

OpenRouter's model-specific Apps view publicly attributes traffic to Hermes Agent and Portkey AI. This is a public adoption signal, not a published implementation study or a KeepRouter customer claim.

Source and verification boundary

Official Qwen3.8 Max 0902 documentation. KeepRouter's live catalog is authoritative for the customer price and enabled endpoint shown here; the maker remains authoritative for upstream model capabilities and limits.

Pricing and implementation guides

Guides

Related models

All models & pricing · Quickstart · KeepRouter vs OpenRouter · Glossary · Get an API key

Catalog facts and prices last changed .

Page content reviewed .