Requests from anywhere on earth flow through one relay and reach DeepSeek, Qwen, GLM, Kimi, and MiniMax — OpenAI-compatible, unified billing, privacy-first defaults.
Works with the official OpenAI client in Python, Node, and every compatible toolchain.
model=deepseek-v4-flash means exactly that. No silent rerouting.
Self-service onboarding with Stripe billing. No China entity, no Alipay, no ICP.
Singapore, Hong Kong, and Tokyo relay paths with dedicated enterprise routes. Just Stripe, one API key, and every frontier model.
Requests ride the lowest-latency path into each upstream, with smart failover across channels when a route degrades.
No plaintext prompts or completions retained. Billing uses metadata only. Optional local redaction via Privacy SDK.
What you request is what you get. No mystery model swaps — every response carries the exact upstream model identifier.
OpenAI-compatible endpoints. Swap models with a single parameter change.
| Model | Context | Input | Output | Latency |
|---|---|---|---|---|
| DeepSeek V4-Flash POPULAR | 1M | $0.264–$0.528/M* | $0.792–$1.584/M* | |
| DeepSeek V4-Pro FLAGSHIP | 1M | $0.792–$1.584/M* | $2.376–$4.752/M* | |
| Qwen Plus NEW | 1M | $0.133/M | $0.800/M | |
| Qwen Max FLAGSHIP | 256K | $0.400/M | $1.600/M | |
| GLM-5.2 NEW | 1M | $1.680/M | $5.280/M | |
| GLM-5.1 STABLE | 203K | $1.164/M | $3.648/M | |
| Kimi K3 NEW | 1M | $3.600/M | $18.00/M | |
| Kimi K2.6 STABLE | 256K | $1.083/M | $4.500/M | |
| MiniMax M3 NEW | 1M | $0.350/M | $1.400/M |
* DeepSeek cache-miss input and output, off-peak to peak; cache-hit input is lower. Includes 20% XinoAPI markup. Updated August 2026. View full pricing →
Self-service onboarding, $2.00 free credits, and 5 provider families through one OpenAI-compatible API.
Self-service · Supported regions only · Privacy SDK available