Changelog
What's new in KeepRouter. Have a request? Email support@keeprouter.com.
2026-08-15
A bilingual AI gateway decision library
- More useful comparisons. The comparison library now covers 15 buyer guides and product comparisons, including fit-based AI gateway and OpenRouter-alternative shortlists, managed versus self-hosted decisions, direct provider APIs, and seven additional gateway or cloud-platform comparisons. Shortlists are organized by operating fit rather than a synthetic ranking.
- Answers and production playbooks. Eight new direct-answer pages explain routing, failover, latency, security, BYOK, gateway selection, and adoption timing. Three new editorial guides provide security, latency-measurement, and failover test plans.
- Stable Simplified Chinese URLs. Every feature, audience, comparison, answer, and blog page now has a crawlable
/zh/...HTML and Markdown twin with localized schema, reciprocalhreflang, and language-preserving internal links. - Clearer reading and discovery. Long-form pages now use answer-first verdicts, sticky contents, decision tables, checklists, step cards, evidence panels, FAQ blocks, and complete related-guide grids. Sitemap,
llms-full.txt, AI index, RSS, SSR, and Markdown derive from the shared registries; rootllms.txtenumerates every English page and links the five Chinese content hubs.
2026-08-13
DeepSeek V4 Pro is generally available
deepseek-v4-pro now resolves to DeepSeek's GA release DeepSeek-V4-Pro-0813. The callable model id has not changed. Its catalog metadata now records the maker-documented 1M-token context window, 384K maximum output, and August 13 release date.
- The current stable DeepSeek ids are
deepseek-v4-proanddeepseek-v4-flash. The retired
deepseek-chat and deepseek-reasoner aliases are no longer suggested by the admin preset.
- DeepSeek now documents Chat Completions, Anthropic-compatible requests, and native Responses API
support for both V4 models. The admin includes separate Chat and Responses channel presets.
- Both V4 models default to thinking mode at high effort; the maker API supports low, high, and max.
2026-07-26
New: DeepSeek V4 Flash — a 1M-token context window
deepseek-v4-flash, a DeepSeek model with a maker-documented 1M-token context window, is now in the production catalog. See DeepSeek V4 Flash pricing & specs.
- Also added:
deepseek-v4-pro, same maker-documented 1M-token context, on the chat route.
Current rates for both are published in Models & Pricing.
2026-07-17
- New: Kimi K3.
kimi-k3— Moonshot AI's flagship with a maker-documented 1M-token (1,048,576) context window — is now on KeepRouter. See Kimi K3 pricing & specs. - More Kimi models. Added
kimi-k2.7-code-highspeed(coding-focused, high-speed serving tier) andkimi-k2.5, joining the existingkimi-k2.6. See Models & Pricing.
2026-07-12
- Public site redesigned. The homepage opens with a working request view backed by the production catalog and health probe. Models and prices use a product directory, and model pages pair facts with a copyable endpoint-specific request.
- Documentation and trust pages reorganized. Quickstart, errors, security, company, comparison, policies, changelog, glossary, and SDK guides use a shared three-column shell with persistent navigation and page contents.
- First-request path corrected. New examples default to
free. Public model CTAs preserve the selected model through email verification, and API key enable and disable actions send boolean state. - Purchase facts moved before signup. The public site now states the minimum card top-up, processing fee formula, tax handling, refund boundary, request-log fields, cache default, and exact health-probe scope.
2026-07-10 · v0.5.0
- Production operations complete. Production and staging are deployed separately; releases run through typecheck, Worker/frontend tests and build, then an isolated staging runtime/schema canary before production. Public staging keeps provider routing and self-serve OTP/email disabled by default rather than copying production credentials.
- Safer operator sessions. The browser admin now uses a signed HttpOnly, Secure, SameSite=Strict cookie; the master admin token remains a non-browser break-glass path and is no longer stored by frontend JavaScript.
- Hardened production sign-in. Transactional email delivery powers OTP, welcome and low-balance mail; a real bot challenge protects the production code-request flow. Staging's official test challenge is not treated as anti-bot protection.
- Production billing. One-time Paddle card top-ups, customer billing portal, signature-verified webhooks, idempotent crediting, and refund/chargeback clawbacks are live.
- Recovery path. Production D1 is exported on a schedule, with Worker rollback and D1 Time Travel documented for operators.
2026-07-05
- New: a free model.
freeis priced at $0 per token. It is intended for testing, prototyping, and evaluating the API before moving to a paid model. See Models & Pricing. - Mistral AI models. Added
mistral-large-latestandmistral-medium-latest; thecodestral-latestcode model and thedevstral-medium-latestcoding-agent model; themagistral-medium-latestandmagistral-small-latestreasoning models; theministral-3b-latest,ministral-8b-latestandministral-14b-latestedge models; and the open-weightopen-mistral-nemo. - Xiaomi MiMo & open-weight models. Added Xiaomi's
mimo-v2.5andmimo-v2.5-pro, plus the open-weightgpt-oss-120b(OpenAI) andgemma-4-31B-it(Google DeepMind). Current rates are published in Models & Pricing.
2026-07-04
- Claude Fable 5 is back.
claude-fable-5is available again on KeepRouter. See Models & Pricing.
2026-06-26
- Transparent top-up processing fee. Top-ups now show the selected credit amount and a separate card-processing fee before payment. Model usage is deducted at the published per-model rates, and the credited balance equals the amount selected.
- Activity log. The Usage page now has a filterable request log (by model & status) with a per-request detail view and CSV export.
- Billing history & invoices. A button on the Credits page opens the Paddle customer portal for invoices, receipts, and payment methods.
- Quickstart guide at /docs/quickstart and an API error reference.
- Low-balance email alerts send a one-time message after a paid balance crosses the warning threshold, plus a welcome email for new accounts.
2026-06-25
- New: Security & Data Handling page. The /security page documents request-log fields, upstream forwarding, and optional response caching. Prompt and completion bodies are not written to the API request log; response caching is off by default.
- Added Doubao models (ByteDance):
doubao-2.0-pro,doubao-2.0-mini,doubao-1.5-pro,doubao-1.5-lite. See Models & Pricing. - Published per-page metadata and structured data, added year-long immutable asset caching, and reduced the page bundle size.
2026-06-24
- Clearer image-model pricing. Image models now show a per-image price (e.g.
$0.04 / image) instead of a token rate. - Full bilingual site. English and 简体中文 are selected from the browser language and can be changed manually.
- New K-monogram logo and social preview cards.
2026-06-23
- Service status page at /status.
- Published Terms of Service, Privacy Policy, Acceptable Use Policy, and Refund Policy.
2026-06-22
- Credit & debit card top-ups. Paddle processes prepaid, pay-as-you-go top-ups as Merchant of Record. There is no monthly subscription.