Simple ways to
start building.
Use the model API at published rates. Reserve dedicated compute with a quote shaped around your workload.
Find your model. See your rate.
Published prices in USD. Token prices below are per one million tokens; other models use their listed billing unit. Check the model's supported options before sending a request.
533 models
| Model | Type | Input / 1M tokens | Output / 1M tokens | Other billing unit |
|---|---|---|---|---|
Mistral: Mixtral 8x22B Instructmistralai/mixtral-8x22b-instruct |
Language | $2.0 | $6.0 | Token based |
Mistral: Sabamistralai/mistral-saba |
Language | $0.2 | $0.6 | Token based |
Mistral: Voxtral Small 24B 2507mistralai/voxtral-small-24b-2507 |
Language | $0.1 | $0.3 | Token based |
MoonshotAI: Kimi K2 0711moonshotai/kimi-k2 |
Language | $0.57 | $2.3 | Token based |
MoonshotAI: Kimi K2 0905moonshotai/kimi-k2-0905 |
Language | $0.6 | $2.5 | Token based |
MoonshotAI: Kimi K2.5moonshotai/kimi-k2.5 |
Language | $0.5 | $2.5 | Token based |
MoonshotAI: Kimi K2.6moonshotai/kimi-k2.6 |
Language | $0.65 | $3.41 | Token based |
MoonshotAI: Kimi K2.7 Codemoonshotai/kimi-k2.7-code |
Language | $0.6712 | $3.35 | Token based |
MoonshotAI: Kimi K2 Thinkingmoonshotai/kimi-k2-thinking |
Language | $0.6 | $2.5 | Token based |
MoonshotAI: Kimi K3moonshotai/kimi-k3 |
Language | $0.8 | $15.0 | Token based |
MoonshotAI: Kimi K3 (batch)moonshotai/kimi-k3:batch |
Language | $2.28 | $11.4 | Token based |
Morph: Morph V3 Fastmorph/morph-v3-fast |
Language | $0.8 | $1.2 | Token based |
Morph: Morph V3 Largemorph/morph-v3-large |
Language | $0.9 | $1.9 | Token based |
MythoMax 13Bgryphe/mythomax-l2-13b |
Language | $0.08 | $0.11 | Token based |
Nex AGI: Nex-N2.5-Mininex-agi/nex-n2.5-mini |
Language | $0.025 | $0.1 | Token based |
Nex AGI: Nex-N2.5-Pronex-agi/nex-n2.5-pro |
Language | $0.075 | $0.25 | Token based |
Nous: Hermes 3 405B Instructnousresearch/hermes-3-llama-3.1-405b |
Language | $1.0 | $1.0 | Token based |
Nous: Hermes 3 70B Instructnousresearch/hermes-3-llama-3.1-70b |
Language | $0.7 | $0.7 | Token based |
Nous: Hermes 4 405Bnousresearch/hermes-4-405b |
Language | $1.0 | $3.0 | Token based |
NVIDIA: Nemotron 3.5 Content Safetynvidia/nemotron-3.5-content-safety |
Language | $0.2 | $0.2 | Token based |
Reported cached input tokens are billed at 10% of the published input rate. Model availability can change. Read the quickstart or open the catalog API (JSON).
Bring your own model.
A dedicated-compute option, custom quoted
Deploy private weights on Actaserve-owned infrastructure, subject to model, licence and runtime qualification. Bring your checkpoint and serving requirements so we can scope model bring-up.
Reserve capacityReady for your first call?
Explore the docs