Simple ways to
start building.
Use the model API at published rates. Reserve dedicated compute with a quote shaped around your workload.
Find your model. See your rate.
Published prices in USD. Token prices below are per one million tokens; other models use their listed billing unit. Check the model's supported options before sending a request.
544 models
| Model | Type | Input / 1M tokens | Output / 1M tokens | Other billing unit |
|---|---|---|---|---|
Google: Gemini 3.5 Flash Litegoogle/gemini-3.5-flash-lite |
Language | $0.3 | $2.5 | Token based |
Google: Gemini 3.5 Flash Lite (batch)google/gemini-3.5-flash-lite:batch |
Language | $0.15 | $1.25 | Token based |
Google: Gemini 3.6 Flashgoogle/gemini-3.6-flash |
Language | $0.75 | $3.75 | Token based |
Google: Gemini 3.6 Flash (batch)google/gemini-3.6-flash:batch |
Language | $0.375 | $1.875 | Token based |
Google: Gemini 3.7 Flashgoogle/gemini-3.7-flash |
Language | $0.75 | $3.75 | Token based |
Google: Gemini 3.7 Flash (batch)google/gemini-3.7-flash:batch |
Language | $0.375 | $1.875 | Token based |
Google: Gemini 3.8 Flashgoogle/gemini-3.8-flash |
Language | $0.75 | $3.75 | Token based |
Google: Gemini 3.8 Flash (batch)google/gemini-3.8-flash:batch |
Language | $0.375 | $1.875 | Token based |
Google: Gemini 3 Flash Previewgoogle/gemini-3-flash-preview |
Language | $0.5 | $3.0 | Token based |
Google: Gemini 3 Flash Preview (batch)google/gemini-3-flash-preview:batch |
Language | $0.25 | $1.5 | Token based |
Google: Gemma 2 27Bgoogle/gemma-2-27b-it |
Language | $0.65 | $0.65 | Token based |
Google: Gemma 3 12Bgoogle/gemma-3-12b-it |
Language | $0.05 | $0.15 | Token based |
Google: Gemma 3 27Bgoogle/gemma-3-27b-it |
Language | $0.08 | $0.45 | Token based |
Google: Gemma 3 4Bgoogle/gemma-3-4b-it |
Language | $0.05 | $0.1 | Token based |
Google: Gemma 4 26B A4B google/gemma-4-26b-a4b-it |
Language | $0.0765 | $0.255 | Token based |
Google: Gemma 4 26B A4B (free)google/gemma-4-26b-a4b-it:free |
Language | $0.0 | $0.0 | Token based |
Google: Gemma 4 31Bgoogle/gemma-4-31b-it |
Language | $0.09 | $0.34 | Token based |
Google: Gemma 4 31B (free)google/gemma-4-31b-it:free |
Language | $0.0 | $0.0 | Token based |
IBM: Granite 4.0 Microibm-granite/granite-4.0-h-micro |
Language | $0.017 | $0.112 | Token based |
IBM: Granite 4.2 8Bibm-granite/granite-4.2-8b |
Language | $0.06 | $0.25 | Token based |
Reported cached input tokens are billed at 10% of the published input rate. Model availability can change. Read the quickstart or open the catalog API (JSON).
Bring your own model.
A dedicated-compute option, custom quoted
Deploy private weights on Actaserve-owned infrastructure, subject to model, licence and runtime qualification. Bring your checkpoint and serving requirements so we can scope model bring-up.
Reserve capacityReady for your first call?
Explore the docs