Changelog

What's new in KeepRouter. Have a request? Email support@keeprouter.com.

2026-08-15

A bilingual AI gateway decision library

  • More useful comparisons. The comparison library now covers 15 buyer guides and product comparisons, including fit-based AI gateway and OpenRouter-alternative shortlists, managed versus self-hosted decisions, direct provider APIs, and seven additional gateway or cloud-platform comparisons. Shortlists are organized by operating fit rather than a synthetic ranking.
  • Answers and production playbooks. Eight new direct-answer pages explain routing, failover, latency, security, BYOK, gateway selection, and adoption timing. Three new editorial guides provide security, latency-measurement, and failover test plans.
  • Stable Simplified Chinese URLs. Every feature, audience, comparison, answer, and blog page now has a crawlable /zh/... HTML and Markdown twin with localized schema, reciprocal hreflang, and language-preserving internal links.
  • Clearer reading and discovery. Long-form pages now use answer-first verdicts, sticky contents, decision tables, checklists, step cards, evidence panels, FAQ blocks, and complete related-guide grids. Sitemap, llms-full.txt, AI index, RSS, SSR, and Markdown derive from the shared registries; root llms.txt enumerates every English page and links the five Chinese content hubs.

2026-08-13

DeepSeek V4 Pro is generally available

deepseek-v4-pro now resolves to DeepSeek's GA release DeepSeek-V4-Pro-0813. The callable model id has not changed. Its catalog metadata now records the maker-documented 1M-token context window, 384K maximum output, and August 13 release date.

  • The current stable DeepSeek ids are deepseek-v4-pro and deepseek-v4-flash. The retired

deepseek-chat and deepseek-reasoner aliases are no longer suggested by the admin preset.

  • DeepSeek now documents Chat Completions, Anthropic-compatible requests, and native Responses API

support for both V4 models. The admin includes separate Chat and Responses channel presets.

  • Both V4 models default to thinking mode at high effort; the maker API supports low, high, and max.

2026-07-26

New: DeepSeek V4 Flash — a 1M-token context window

deepseek-v4-flash, a DeepSeek model with a maker-documented 1M-token context window, is now in the production catalog. See DeepSeek V4 Flash pricing & specs.

  • Also added: deepseek-v4-pro, same maker-documented 1M-token context, on the chat route.

Current rates for both are published in Models & Pricing.

2026-07-17

  • New: Kimi K3. kimi-k3 — Moonshot AI's flagship with a maker-documented 1M-token (1,048,576) context window — is now on KeepRouter. See Kimi K3 pricing & specs.
  • More Kimi models. Added kimi-k2.7-code-highspeed (coding-focused, high-speed serving tier) and kimi-k2.5, joining the existing kimi-k2.6. See Models & Pricing.

2026-07-12

  • Public site redesigned. The homepage opens with a working request view backed by the production catalog and health probe. Models and prices use a product directory, and model pages pair facts with a copyable endpoint-specific request.
  • Documentation and trust pages reorganized. Quickstart, errors, security, company, comparison, policies, changelog, glossary, and SDK guides use a shared three-column shell with persistent navigation and page contents.
  • First-request path corrected. New examples default to free. Public model CTAs preserve the selected model through email verification, and API key enable and disable actions send boolean state.
  • Purchase facts moved before signup. The public site now states the minimum card top-up, processing fee formula, tax handling, refund boundary, request-log fields, cache default, and exact health-probe scope.

2026-07-10 · v0.5.0

  • Production operations complete. Production and staging are deployed separately; releases run through typecheck, Worker/frontend tests and build, then an isolated staging runtime/schema canary before production. Public staging keeps provider routing and self-serve OTP/email disabled by default rather than copying production credentials.
  • Safer operator sessions. The browser admin now uses a signed HttpOnly, Secure, SameSite=Strict cookie; the master admin token remains a non-browser break-glass path and is no longer stored by frontend JavaScript.
  • Hardened production sign-in. Transactional email delivery powers OTP, welcome and low-balance mail; a real bot challenge protects the production code-request flow. Staging's official test challenge is not treated as anti-bot protection.
  • Production billing. One-time Paddle card top-ups, customer billing portal, signature-verified webhooks, idempotent crediting, and refund/chargeback clawbacks are live.
  • Recovery path. Production D1 is exported on a schedule, with Worker rollback and D1 Time Travel documented for operators.

2026-07-05

  • New: a free model. free is priced at $0 per token. It is intended for testing, prototyping, and evaluating the API before moving to a paid model. See Models & Pricing.
  • Mistral AI models. Added mistral-large-latest and mistral-medium-latest; the codestral-latest code model and the devstral-medium-latest coding-agent model; the magistral-medium-latest and magistral-small-latest reasoning models; the ministral-3b-latest, ministral-8b-latest and ministral-14b-latest edge models; and the open-weight open-mistral-nemo.
  • Xiaomi MiMo & open-weight models. Added Xiaomi's mimo-v2.5 and mimo-v2.5-pro, plus the open-weight gpt-oss-120b (OpenAI) and gemma-4-31B-it (Google DeepMind). Current rates are published in Models & Pricing.

2026-07-04

  • Claude Fable 5 is back. claude-fable-5 is available again on KeepRouter. See Models & Pricing.

2026-06-26

  • Transparent top-up processing fee. Top-ups now show the selected credit amount and a separate card-processing fee before payment. Model usage is deducted at the published per-model rates, and the credited balance equals the amount selected.
  • Activity log. The Usage page now has a filterable request log (by model & status) with a per-request detail view and CSV export.
  • Billing history & invoices. A button on the Credits page opens the Paddle customer portal for invoices, receipts, and payment methods.
  • Quickstart guide at /docs/quickstart and an API error reference.
  • Low-balance email alerts send a one-time message after a paid balance crosses the warning threshold, plus a welcome email for new accounts.

2026-06-25

  • New: Security & Data Handling page. The /security page documents request-log fields, upstream forwarding, and optional response caching. Prompt and completion bodies are not written to the API request log; response caching is off by default.
  • Added Doubao models (ByteDance): doubao-2.0-pro, doubao-2.0-mini, doubao-1.5-pro, doubao-1.5-lite. See Models & Pricing.
  • Published per-page metadata and structured data, added year-long immutable asset caching, and reduced the page bundle size.

2026-06-24

  • Clearer image-model pricing. Image models now show a per-image price (e.g. $0.04 / image) instead of a token rate.
  • Full bilingual site. English and 简体中文 are selected from the browser language and can be changed manually.
  • New K-monogram logo and social preview cards.

2026-06-23

  • Service status page at /status.
  • Published Terms of Service, Privacy Policy, Acceptable Use Policy, and Refund Policy.

2026-06-22

  • Credit & debit card top-ups. Paddle processes prepaid, pay-as-you-go top-ups as Merchant of Record. There is no monthly subscription.