Simple ways to
start building.
Use the model API at published rates. Reserve dedicated compute with a quote shaped around your workload.
Find your model. See your rate.
Published prices in USD. Token prices below are per one million tokens; other models use their listed billing unit. Check the model's supported options before sending a request.
533 models
| Model | Type | Input / 1M tokens | Output / 1M tokens | Other billing unit |
|---|---|---|---|---|
Z.ai: GLM 5.1z-ai/glm-5.1 |
Language | $0.966 | $3.036 | Token based |
Z.ai: GLM 5.2z-ai/glm-5.2 |
Language | $0.06 | $7.0 | Token based |
Z.ai: GLM 5.3z-ai/glm-5.3 |
Language | $0.039 | $4.8 | Token based |
Z.ai: GLM 5.3 (batch)z-ai/glm-5.3:batch |
Language | $0.45 | $2.0 | Token based |
Z.ai: GLM 5.3 Flashz-ai/glm-5.3-flash |
Language | $0.15 | $0.5 | Token based |
Z.ai: GLM 5.3 Flash (batch)z-ai/glm-5.3-flash:batch |
Language | $0.06 | $0.2 | Token based |
Z.ai: GLM 5.3 FlashXz-ai/glm-5.3-flashx |
Language | $0.37 | $1.25 | Token based |
Z.ai: GLM 5.3 Primez-ai/glm-5.3-prime |
Language | $2.8 | $8.8 | Token based |
Z.ai: GLM 5 Turboz-ai/glm-5-turbo |
Language | $1.2 | $4.0 | Token based |
Z.ai: GLM 5V Turboz-ai/glm-5v-turbo |
Language | $1.2 | $4.0 | Token based |
Grok Imagine Videoxai/grok-imagine-video/text-to-video |
Video | Not token based | Not token based | $0.05 / second |
Kling 1.6 Standardkwai/kling-1.6-standard |
Video | Not token based | Not token based | $0.045 / second |
PixVerse v6fal-ai/pixverse/v6/text-to-video |
Video | Not token based | Not token based | $0.005 / second |
Reported cached input tokens are billed at 10% of the published input rate. Model availability can change. Read the quickstart or open the catalog API (JSON).
Bring your own model.
A dedicated-compute option, custom quoted
Deploy private weights on Actaserve-owned infrastructure, subject to model, licence and runtime qualification. Bring your checkpoint and serving requirements so we can scope model bring-up.
Reserve capacityReady for your first call?
Explore the docs