Kimi K3 and Kimi K3 Fast with ZDR and US-based providers now on AI Gateway
Source
vercel.com
Author
Jerilyn Zheng
Date
Why it matters
If you want to build on Moonshot's Kimi K3 but need US-based inferenceRunning a trained model to get answers — the phase where AI is actually used, as opposed to trained.Full definition → and no data retention, the AI Gateway now gives you that with a couple of provider options (speed, inferenceRegion, zeroDataRetention) and automatic failover across Baseten/Fireworks — no code changes or per-provider contracts.
Terms in this piece · Glossary
inference — Running a trained model to get answers — the phase where AI is actually used, as opposed to trained.
Key quotes
“Running Kimi K3 on US-based providers lets teams with data residency and compliance requirements use the model on US infrastructure.”
“Because AI Gateway serves the models from multiple providers, it automatically routes across them for failover, higher uptime, and more available throughput than any single provider offers.”
“Kimi K3 Fast trades a higher per-token cost for lower latency.”