Simple ways to
start building.
Use the model API at published rates. Reserve dedicated compute with a quote shaped around your workload.
Find your model. See your rate.
Published prices in USD. Token prices below are per one million tokens; other models use their listed billing unit. Check the model's supported options before sending a request.
533 models
| Model | Type | Input / 1M tokens | Output / 1M tokens | Other billing unit |
|---|---|---|---|---|
Thinking Machines: Inkling Small (free)thinkingmachines/inkling-small:free |
Language | $0.0 | $0.0 | Token based |
Upstage: Solar Mini 4upstage/solar-mini4 |
Language | $0.05 | $0.2 | Token based |
Upstage: Solar Pro 3upstage/solar-pro-3 |
Language | $0.15 | $0.6 | Token based |
Upstage: Solar Pro 4upstage/solar-pro4 |
Language | $0.09 | $0.36 | Token based |
Venice: Uncensoredcognitivecomputations/dolphin-mistral-24b-venice-edition |
Language | $0.2 | $0.9 | Token based |
WizardLM-2 8x22Bmicrosoft/wizardlm-2-8x22b |
Language | $0.62 | $0.62 | Token based |
Writer: Palmyra X5writer/palmyra-x5 |
Language | $0.6 | $6.0 | Token based |
Xiaomi: MiMo-V2.5xiaomi/mimo-v2.5 |
Language | $0.14 | $0.28 | Token based |
Xiaomi: MiMo-V2.5-Proxiaomi/mimo-v2.5-pro |
Language | $0.435 | $0.87 | Token based |
Xiaomi: MiMo-V2.6-Flashxiaomi/mimo-v2.6-flash |
Language | $0.14 | $0.28 | Token based |
Xiaomi: MiMo-V2.6-Proxiaomi/mimo-v2.6-pro |
Language | $0.435 | $0.87 | Token based |
Xiaomi: MiMo-V2.6-Pro-UltraSpeedxiaomi/mimo-v2.6-pro-ultraspeed |
Language | $4.35 | $8.7 | Token based |
Z.ai: GLM 4.5z-ai/glm-4.5 |
Language | $0.6 | $2.2 | Token based |
Z.ai: GLM 4.5 Airz-ai/glm-4.5-air |
Language | $0.13 | $0.85 | Token based |
Z.ai: GLM 4.5Vz-ai/glm-4.5v |
Language | $0.6 | $1.8 | Token based |
Z.ai: GLM 4.6z-ai/glm-4.6 |
Language | $0.5 | $2.0 | Token based |
Z.ai: GLM 4.6Vz-ai/glm-4.6v |
Language | $0.3 | $0.9 | Token based |
Z.ai: GLM 4.7z-ai/glm-4.7 |
Language | $0.6 | $2.2 | Token based |
Z.ai: GLM 4.7 Flashz-ai/glm-4.7-flash |
Language | $0.0605 | $0.4 | Token based |
Z.ai: GLM 5z-ai/glm-5 |
Language | $0.6 | $1.92 | Token based |
Reported cached input tokens are billed at 10% of the published input rate. Model availability can change. Read the quickstart or open the catalog API (JSON).
Bring your own model.
A dedicated-compute option, custom quoted
Deploy private weights on Actaserve-owned infrastructure, subject to model, licence and runtime qualification. Bring your checkpoint and serving requirements so we can scope model bring-up.
Reserve capacityReady for your first call?
Explore the docs