GPT-56T 861 —
MUSE-SPK 835 -0.7%
GPT-56SC 828 -5.2%
QWEN-38X 824 —
CL-OP55X 822 —
GROK-46H 822 -5%
GPT-6A 820 —
GLM-5 784 -8.4%
CL-FAB5H 743 -5.6%
KIMI-K3X 742 -8.4%
CL-OP5H 720 -5.8%
CL-OP5X 709 -18%
CL-OP46H 698 -5.9%
CL-OP47H 690 -5.9%
GEM-38FH 677 +0.1%
GEM-37FH 657 -24%
GPT-56S 622 —
CL-OP47 582 -0.7%
GPT-55H 582 —
INKL 531 —
GEM-31P 513 —
GEM-3P 499 —
CL-OP46 496 -0.2%
CL-OP48 490 —
GPT-56T 861 —
MUSE-SPK 835 -0.7%
GPT-56SC 828 -5.2%
QWEN-38X 824 —
CL-OP55X 822 —
GROK-46H 822 -5%
GPT-6A 820 —
GLM-5 784 -8.4%
CL-FAB5H 743 -5.6%
KIMI-K3X 742 -8.4%
CL-OP5H 720 -5.8%
CL-OP5X 709 -18%
CL-OP46H 698 -5.9%
CL-OP47H 690 -5.9%
GEM-38FH 677 +0.1%
GEM-37FH 657 -24%
GPT-56S 622 —
CL-OP47 582 -0.7%
GPT-55H 582 —
INKL 531 —
GEM-31P 513 —
GEM-3P 499 —
CL-OP46 496 -0.2%
CL-OP48 490 —
← Back to feed

OpenRouter US In-Region Routing: DeepSeek V4 Pro and Kimi K3 Now Processable Inside American Data Centers

OpenRouter’s US In-Region Routing is now live, joining the EU tier that launched earlier this year. Requests sent to us.openrouter.ai are decrypted and served only by providers running in the United States — the originating lab is not involved in request handling.

What It Resolves

Chinese open-weight models have faced a consistent procurement problem at US enterprises: even when the weights are legally accessible and technically useful, routing inference traffic through Chinese infrastructure is a non-starter for legal, compliance, and IT security teams. In-Region Routing addresses that directly.

Three Chinese-origin models are currently available on the US tier:

  • DeepSeek V4 Pro — served by Baseten and Fireworks from US data centers
  • Kimi K3 — Moonshot’s 2.8T open-weight model with 93.4% SWE-bench Verified
  • GLM 5.2 — also available on the EU tier via Mistral’s European infrastructure

When a request goes through us.openrouter.ai, it hits whichever US-based provider hosts the model. DeepSeek’s servers in China are not in the data path. For GLM 5.2, EU routing routes through Mistral.

Access and Pricing

In-Region Routing is gated to Business and Enterprise plans. Standard OpenRouter per-token pricing applies — there is no premium for in-region routing. The endpoint change is the only integration difference:

https://us.openrouter.ai/api/v1   # US in-region
https://eu.openrouter.ai/api/v1   # EU in-region

Model selection on the in-region endpoint is limited to models with confirmed in-region hosting. Not every model on the standard endpoint is available in-region — organizations need to verify the in-region model list before migrating existing integrations.

The Open-Weight Angle

OpenRouter’s own traffic data shows open-weight models have grown as a share of total tokens for both US and EU originating requests. The growth is partly driven by cost — DeepSeek V4 Pro and Kimi K3 carry significantly lower per-token prices than frontier proprietary models — and partly by the emergence of Chinese labs producing competitive models.

In-Region Routing makes that cost case actionable for a category of enterprises that previously had to self-host to achieve the same data residency guarantee. Self-hosting Kimi K3’s 2.8T parameter weights requires non-trivial GPU infrastructure; routing through a US-hosted Fireworks or Baseten endpoint removes that barrier.

NVIDIA’s Nemotron 3 Ultra and Thinking Machines’ Inkling are also included in the in-region model set — US-lab open-weight models that have gained commercial traction but until now lacked a managed residency path on the OpenRouter platform.

The Stripe Context

OpenRouter announced it was joining Stripe in August. The in-region routing launch is the first major product push under that arrangement. Stripe’s enterprise sales surface and compliance infrastructure are natural distribution channels for features that reduce friction for regulated-industry customers, which is the primary market for data residency routing.