GLM 5.3 is Live on EmberCloud. Try It Today
New: OpenAI-compatible chat + embeddings

Affordable tokens at blazing fast speeds

Serverless GPU inference for open source models with predictable latency, simple pricing, and drop-in OpenAI APIs.

Zero cold starts Usage + rate limits OpenAI compatible
Global Infrastructure

Worldwide reach,
blazing fast inference

Distributed cloud infrastructure with an inference engine built for speed.

Integrate in seconds

Our API is fully compatible with the OpenAI SDK. Simply change the base URL and API key to switch to open-source models.

1

Get your API Key

Sign up and generate a key in the dashboard.

2

Configure Client

Point your existing SDK to EmberCloud endpoints.

main.py
from openai import OpenAI

client = OpenAI(
    base_url="https://api.embercloud.ai/v1",
    api_key="ember_sk_..."
)

completion = client.chat.completions.create(
    model="glm-4.7",
    messages=[
        {"role": "user", "content": "Hello World!"}
    ]
)

print(completion.choices[0].message.content)
Model Library

Production-ready models

Access top-tier open models through a unified, OpenAI-compatible API.

New
Chat
GLM

GLM-5.3

1.3M ctx
$0.756 in · $2.38 out / 1M
View pricing
Fast
Fast
GLM

GLM-5.3 Flash

1.3M ctx
$0.135 in · $0.450 out / 1M
View pricing
Chat
GLM

GLM 5.2

203K ctx
$1.26 in · $3.96 out / 1M
View pricing
Chat
GLM

GLM 5.1

203K ctx
$0.931 in · $2.93 out / 1M
View pricing
Chat
GLM

GLM 5

745B MoE 203K ctx
$0.570 in · $1.82 out / 1M
View pricing
Chat
GLM

GLM 4.7

355B MoE 200K ctx
$0.380 in · $1.66 out / 1M
View pricing
Fast
Fast
GLM

GLM 4.7 Flash

30B MoE 200K ctx
$0.057 in · $0.380 out / 1M
View pricing
Popular
Code
GLM

GLM 4.6

200K ctx
$0.380 in · $1.55 out / 1M
View pricing
Chat
GLM

GLM 4.5

355B MoE 131K ctx
$0.570 in · $2.09 out / 1M
View pricing
Value
Chat
GLM

GLM 4.5 Air

131K ctx
$0.123 in · $0.807 out / 1M
View pricing
New
Chat
GLM

GLM-5.3

1.3M ctx
$0.756 in · $2.38 out / 1M
View pricing
Fast
Fast
GLM

GLM-5.3 Flash

1.3M ctx
$0.135 in · $0.450 out / 1M
View pricing
Chat
GLM

GLM 5.2

203K ctx
$1.26 in · $3.96 out / 1M
View pricing
Chat
GLM

GLM 5.1

203K ctx
$0.931 in · $2.93 out / 1M
View pricing
Chat
GLM

GLM 5

745B MoE 203K ctx
$0.570 in · $1.82 out / 1M
View pricing
Chat
GLM

GLM 4.7

355B MoE 200K ctx
$0.380 in · $1.66 out / 1M
View pricing
Fast
Fast
GLM

GLM 4.7 Flash

30B MoE 200K ctx
$0.057 in · $0.380 out / 1M
View pricing
Popular
Code
GLM

GLM 4.6

200K ctx
$0.380 in · $1.55 out / 1M
View pricing
Chat
GLM

GLM 4.5

355B MoE 131K ctx
$0.570 in · $2.09 out / 1M
View pricing
Value
Chat
GLM

GLM 4.5 Air

131K ctx
$0.123 in · $0.807 out / 1M
View pricing
New
Chat
Kimi

Kimi K3

1M ctx
$2.70 in · $13.50 out / 1M
View pricing
Code
Kimi

Kimi K2.7 Code

262K ctx
$0.635 in · $2.97 out / 1M
View pricing
Chat
Kimi

Kimi K2.6

262K ctx
$0.855 in · $3.60 out / 1M
View pricing
Chat
Kimi

Kimi K2.5

262K ctx
$0.356 in · $1.98 out / 1M
View pricing
New
Chat
MiniMax

MiniMax M3

1M ctx
$0.270 in · $1.08 out / 1M
View pricing
Chat
MiniMax

MiniMax M2.7

205K ctx
$0.270 in · $1.08 out / 1M
View pricing
Chat
MiniMax

MiniMax M2.5

196K ctx
$0.143 in · $0.855 out / 1M
View pricing
Flagship
Chat
DeepSeek

DeepSeek V4 Pro

1M ctx
$0.851 in · $1.70 out / 1M
View pricing
Fast
Fast
DeepSeek

DeepSeek V4 Flash

1M ctx
$0.065 in · $0.129 out / 1M
View pricing
Value
Code
Qwen

Qwen3 Coder Next

262K ctx
$0.108 in · $0.675 out / 1M
View pricing
New
Chat
Kimi

Kimi K3

1M ctx
$2.70 in · $13.50 out / 1M
View pricing
Code
Kimi

Kimi K2.7 Code

262K ctx
$0.635 in · $2.97 out / 1M
View pricing
Chat
Kimi

Kimi K2.6

262K ctx
$0.855 in · $3.60 out / 1M
View pricing
Chat
Kimi

Kimi K2.5

262K ctx
$0.356 in · $1.98 out / 1M
View pricing
New
Chat
MiniMax

MiniMax M3

1M ctx
$0.270 in · $1.08 out / 1M
View pricing
Chat
MiniMax

MiniMax M2.7

205K ctx
$0.270 in · $1.08 out / 1M
View pricing
Chat
MiniMax

MiniMax M2.5

196K ctx
$0.143 in · $0.855 out / 1M
View pricing
Flagship
Chat
DeepSeek

DeepSeek V4 Pro

1M ctx
$0.851 in · $1.70 out / 1M
View pricing
Fast
Fast
DeepSeek

DeepSeek V4 Flash

1M ctx
$0.065 in · $0.129 out / 1M
View pricing
Value
Code
Qwen

Qwen3 Coder Next

262K ctx
$0.108 in · $0.675 out / 1M
View pricing
Transparent Pricing

Flexible token pricing

Pay only for what you generate. No idle costs.

ModelContextInputOutputCached Input
GLMGLM-5.3New
1.3M$0.756 / 1M$2.38 / 1M$0.140 / 1M
GLMGLM-5.3 FlashFast
1.3M$0.135 / 1M$0.450 / 1M$0.045 / 1M
GLMGLM 5.2
203K$1.26 / 1M$3.96 / 1M$0.234 / 1M
GLMGLM 5.1
203K$0.931 / 1M$2.93 / 1M$0.173 / 1M
GLMGLM 5
203K$0.570 / 1M$1.82 / 1M$0.114 / 1M
GLMGLM 4.7
200K$0.380 / 1M$1.66 / 1M$0.076 / 1M
GLMGLM 4.7 FlashFast
200K$0.057 / 1M$0.380 / 1M$0.0095 / 1M
GLMGLM 4.6Popular
200K$0.380 / 1M$1.55 / 1M$0.070 / 1M
GLMGLM 4.5
131K$0.570 / 1M$2.09 / 1M$0.105 / 1M
GLMGLM 4.5 AirValue
131K$0.123 / 1M$0.807 / 1M$0.024 / 1M
KimiKimi K3New
1M$2.70 / 1M$13.50 / 1M$0.270 / 1M
KimiKimi K2.7 Code
262K$0.635 / 1M$2.97 / 1M$0.162 / 1M
KimiKimi K2.6
262K$0.855 / 1M$3.60 / 1M$0.144 / 1M
KimiKimi K2.5
262K$0.356 / 1M$1.98 / 1M$0.225 / 1M
MiniMaxMiniMax M3New
1M$0.270 / 1M$1.08 / 1M$0.054 / 1M
MiniMaxMiniMax M2.7
205K$0.270 / 1M$1.08 / 1M$0.054 / 1M
MiniMaxMiniMax M2.5
196K$0.143 / 1M$0.855 / 1M$0.040 / 1M
DeepSeekDeepSeek V4 ProFlagship
1M$0.851 / 1M$1.70 / 1M$0.071 / 1M
DeepSeekDeepSeek V4 FlashFast
1M$0.065 / 1M$0.129 / 1M$0.013 / 1M
QwenQwen3 Coder NextValue
262K$0.108 / 1M$0.675 / 1M$0.060 / 1M