Kimi K3
Up to 1M context with native vision and configurable reasoning for coding, knowledge work and agent tasks.
Provider documentationOne OpenAI-compatible API for China and global AI models.
Built for global developers who need broader model choice, reliable routing and simple usage-based billing.
Broader model choice, one production-ready API
Reliable routing for responsive model calls across leading AI providers and production workloads.
Clear, usage-based billing makes it easy to test, scale and manage model costs.
Use leading China models including Kimi, GLM, HappyHorse, Qwen, Wan, DeepSeek and BytePlus, alongside GPT, Claude and Gemini.
Multi-node routing is continually tuned for responsive, reliable API access across regions.
OpenAI-compatible endpoints: usually only change base_url in existing code.
Technical and business support for teams building and scaling with AI.
One API key for leading China models and global model choice
Up to 1M context with native vision and configurable reasoning for coding, knowledge work and agent tasks.
Provider documentation1M context, up to 128K output, plus function calling and structured output for long-running development workflows.
Provider documentation1M context with reasoning, function calling and built-in tools. This model remains in preview.
Provider documentation1M context, up to 384K output, tool calls and thinking modes for agentic and reasoning workloads.
Provider documentationImage and video generation options for creative media workflows, alongside text and reasoning models.
Provider documentationSupported models
Availability
Avg latency
Chinese support
Provider reference prices with the billing unit shown for each model
| Models | Input | Output |
|---|---|---|
| Leading China Models | ||
| GLM-5.2glm-5.2 | CNY 8CNY 2 cache hit | CNY 28 |
| Wan 2.7 T2Vwan2.7-t2v | N/AOutput-only billing | CNY 0.733924/s720P international · 1080P: CNY 1.100886/s |
| HappyHorse 1.1 T2Vhappyhorse-1.1-t2v | N/AOutput-only billing | CNY 0.90/s720P global list price · 1080P: CNY 1.20/s |
| Seedream 5.0 Prodola-seedream-5-0-pro | Free first imageThen USD 0.003/image | USD 0.045/imageUp to 2.36M pixels · USD 0.09 above |
| Seed 2.1 Turbodola-seed-2-1-turbo | USD 0.50USD 0.10 cache hit | USD 2.50 |
| Qwen3.8 Max Previewqwen3.8-max-preview | Token Plan onlyNo PAYG token rate published | See plan |
| Kimi K3kimi-k3 | CNY 20CNY 2 cache hit | CNY 100 |
| DeepSeek V4 Prodeepseek-v4-pro | USD 0.435USD 0.003625 cache hit | USD 0.87 |
| Models | Input | Output |
|---|---|---|
| Latest Models | ||
| Claude Sonnet 5claude-sonnet-5 | $2introductory through Aug 31, 2026 | $10 |
| Claude Haiku 4.5claude-haiku-4-5 | $1 | $5 |
| Models | Input | Output |
|---|---|---|
| Claude Opus 4.6claude-opus-4-6 | $5 | $25 |
| Claude Sonnet 4.5claude-sonnet-4-5 | $3 | $15 |
| Claude Opus 4.5claude-opus-4-5 | $5 | $25 |
| Claude Opus 4.1claude-opus-4-1 | $15 | $75 |
| Claude Sonnet 4claude-sonnet-4 | $3 | $15 |
| Claude Haiku 3claude-3-haiku | $0.25 | $1.25 |
| Models | Input | Output |
|---|---|---|
| Gemini 3 Series | ||
| Gemini 3.5 Flashgemini-3.5-flash | See official pricingStable model for agentic and coding tasks | See official pricing |
| Gemini 3.1 Progemini-3.1-pro-preview | $2> 200k input tokens: $4Input: text / image / video / audio | $12> 200k input tokens: $18Text output |
| Gemini 3.1 Flash Imagegemini-3.1-flash-image-preview | $0.50Input: text / image | $3Image output: $60 |
| Gemini 3.1 Flash-Litegemini-3.1-flash-lite-preview | $0.25Audio input: $0.50 | $1.50 |
| Gemini 3 Progemini-3-pro-preview | $2> 200k input tokens: $4 | $12> 200k input tokens: $18 |
| Gemini 3 Flashgemini-3-flash-preview | $0.50Audio input: $1 | $3 |
| Models | Input | Output |
|---|---|---|
| Gemini 3 Pro Imagegemini-3-pro-image-preview | $2Input: text / image | $12Image output: $120 |
| Gemini 2.5 Pro Computer Usegemini-2.5-pro-computer-use-preview-1025 | $1.25> 200k input tokens: $2.50 | $10> 200k input tokens: $15 |
| Gemini 2.5 Flash Imagegemini-2.5-flash-image | $0.30Input: text / image | $2.50Image output: $30 |
| Gemini 2.5 Flash Live APIgemini-2.5-flash-live | Text $0.5Audio $3 / video and image $3 | Text $2Audio $12 |
| Models | Input | Output |
|---|---|---|
| Gemini 2.5 Progemini-2.5-pro | $1.25> 200k input tokens: $2.50 | $10> 200k input tokens: $15 |
| Gemini 2.5 Flashgemini-2.5-flash | $0.30Audio input: $1 | $2.50 |
| Gemini 2.5 Flash Litegemini-2.5-flash-lite | $0.10Audio input: $0.30 | $0.40 |
| Models | Input | Output |
|---|---|---|
| Latest Flagship Models | ||
| GPT-5.6Sol / Terra / Luna | See official pricingThree model tiers | See official pricing |
| Models | Input | Output |
|---|---|---|
| GPT-5.5gpt-5.5 | $5.00 | $30.00 |
| GPT-5.4gpt-5.4 | $2.50Long > 272k: $5.00 | $15.00Long > 272k: $22.50 |
| GPT-5.2gpt-5.2 | $1.25 | $10.00 |
| GPT-5gpt-5 | $1.25 | $10.00 |
| GPT-5 Minigpt-5-mini | $0.25 | $2.00 |
| GPT-5 Nanogpt-5-nano | $0.05 | $0.40 |
| GPT-4.1gpt-4.1 | $2.00 | $8.00 |
| GPT-4ogpt-4o | $2.50 | $10.00 |
| GPT-4o Minigpt-4o-mini | $0.15 | $0.60 |
| Models | Input | Output |
|---|---|---|
| GPT-Realtime-2gpt-realtime-2 | Audio $32.00cached audio $0.40 | Audio $64.00 |
| GPT-Realtime 1.5gpt-realtime-1.5 | Audio $32.00Text $4.00 / Image $5.00 | Audio $64.00Text $16.00 |
| GPT-Realtime Minigpt-realtime-mini | Audio $10.00Text $0.60 / Image $0.80 | Audio $20.00Text $2.40 |
| GPT-Audiogpt-audio | Audio $32.00Text $2.50 | Audio $64.00Text $10.00 |
| Models | Input | Output |
|---|---|---|
| ChatGPT Images 2.0chatgpt-image-latest | Image $8.00Text $5.00 | Image $32.00Text $10.00 |
| GPT-Image 1.5gpt-image-1.5 | Image $8.00Text $5.00 | Image $32.00Text $10.00 |
| GPT-Image 1 Minigpt-image-1-mini | Image $2.50Text $2.00 | Image $8.00 |
| Transcription Models | ||
| GPT-4o Transcribegpt-4o-transcribe | $2.50 | $10.00 |
| GPT-4o Mini Transcribegpt-4o-mini-transcribe | $1.25 | $5.00 |
| O Series | ||
| o4-minio4-mini | $1.10 | $4.40 |
| o3o3 | $2.00 | $8.00 |
| o3-proo3-pro | $20.00 | $80.00 |
Integrate in four steps and start using the API right away
Register and verify your account
Alipay / WeChat Pay
Create an API key in one click
Change base_url and call
pip install openai # Set API key export ONEAI_API_KEY="your-api-key"
import os from openai import OpenAI client = OpenAI( api_key=os.environ["ONEAI_API_KEY"], base_url="https://uri/v1", ) response = client.responses.create( model="gpt-5.6-sol", input="Say hello in one sentence.", stream=False, ) print(response)
Everything you need to get started
Get an API key and start building with China and global AI models