Simple ways to
start building.
Use the model API at published rates. Reserve dedicated compute with a quote shaped around your workload.
Find your model. See your rate.
Published prices in USD. Token prices below are per one million tokens; other models use their listed billing unit. Check the model's supported options before sending a request.
544 models
| Model | Type | Input / 1M tokens | Output / 1M tokens | Other billing unit |
|---|---|---|---|---|
Inception: Mercury 2inception/mercury-2 |
Language | $0.25 | $0.75 | Token based |
Inception: Mercury 2.5inception/mercury-2.5 |
Language | $0.04 | $0.15 | Token based |
inclusionAI: Ling 3.0 Flashinclusionai/ling-3.0-flash |
Language | $0.021 | $0.063 | Token based |
inclusionAI: Ling 3.0 Flash Fininclusionai/ling-3.0-flash-fin |
Language | $0.042 | $0.1232 | Token based |
inclusionAI: Ling 3.0 Flash Santeinclusionai/ling-3.0-flash-sante |
Language | $0.042 | $0.1232 | Token based |
inclusionAI: Ling 3.0 Flash VLinclusionai/ling-3.0-flash-vl |
Language | $0.021 | $0.0616 | Token based |
inclusionAI: Ling 3.1 Flashinclusionai/ling-3.1-flash |
Language | $0.0 | $0.0 | Token based |
Inference.net: Schematron V2 Smallinference-net/schematron-v2-small |
Language | $0.05 | $0.23 | Token based |
Inference.net: Schematron V2 Turboinference-net/schematron-v2-turbo |
Language | $0.03 | $0.15 | Token based |
LiquidAI: LFM2.5-2.6B (free)liquid/lfm-2.5-2.6b:free |
Language | $0.0 | $0.0 | Token based |
Llama 3.1 8Bmeta/llama-3.1-8b |
Language | $0.02 | $0.03 | Token based |
Magnum v4 72Banthracite-org/magnum-v4-72b |
Language | $2.5 | $5.0 | Token based |
Mancer: Weaver (alpha)mancer/weaver |
Language | $0.4 | $0.75 | Token based |
Meituan: LongCat 2.0meituan/longcat-2.0 |
Language | $0.3 | $1.2 | Token based |
Meta: Llama 3.1 70B Instructmeta-llama/llama-3.1-70b-instruct |
Language | $0.4 | $0.4 | Token based |
Meta: Llama 3.1 8B Instructmeta-llama/llama-3.1-8b-instruct |
Language | $0.05 | $0.08 | Token based |
Meta: Llama 3.2 1B Instructmeta-llama/llama-3.2-1b-instruct |
Language | $0.027 | $0.201 | Token based |
Meta: Llama 3.2 3B Instructmeta-llama/llama-3.2-3b-instruct |
Language | $0.05 | $0.33 | Token based |
Meta: Llama 3.3 70B Instructmeta-llama/llama-3.3-70b-instruct |
Language | $0.1 | $0.32 | Token based |
Meta: Llama 4 Maverickmeta-llama/llama-4-maverick |
Language | $0.1875 | $0.6525 | Token based |
Reported cached input tokens are billed at 10% of the published input rate. Model availability can change. Read the quickstart or open the catalog API (JSON).
Bring your own model.
A dedicated-compute option, custom quoted
Deploy private weights on Actaserve-owned infrastructure, subject to model, licence and runtime qualification. Bring your checkpoint and serving requirements so we can scope model bring-up.
Reserve capacityReady for your first call?
Explore the docs