Simple ways to
start building.
Use the model API at published rates. Reserve dedicated compute with a quote shaped around your workload.
Find your model. See your rate.
Published prices in USD. Token prices below are per one million tokens; other models use their listed billing unit. Check the model's supported options before sending a request.
543 models
| Model | Type | Input / 1M tokens | Output / 1M tokens | Other billing unit |
|---|---|---|---|---|
Mistral: Codestral Embed 2505mistralai/codestral-embed-2505 |
Embeddings | $0.15 | $0.0 | Token based |
Mistral: Mistral Embed 2312mistralai/mistral-embed-2312 |
Embeddings | $0.1 | $0.0 | Token based |
NVIDIA: Llama Nemotron Embed VL 1B V2 (free)nvidia/llama-nemotron-embed-vl-1b-v2:free |
Embeddings | $0.0 | $0.0 | Token based |
NVIDIA: Nemotron 3 Embed 1B (free)nvidia/nemotron-3-embed-1b:free |
Embeddings | $0.0 | $0.0 | Token based |
OpenAI: Text Embedding 3 Largeopenai/text-embedding-3-large |
Embeddings | $0.13 | $0.0 | Token based |
OpenAI: Text Embedding 3 Smallopenai/text-embedding-3-small |
Embeddings | $0.02 | $0.0 | Token based |
OpenAI: Text Embedding Ada 002openai/text-embedding-ada-002 |
Embeddings | $0.1 | $0.0 | Token based |
Perplexity: Embed V1 0.6Bperplexity/pplx-embed-v1-0.6b |
Embeddings | $0.004 | $0.0 | Token based |
Perplexity: Embed V1 4Bperplexity/pplx-embed-v1-4b |
Embeddings | $0.03 | $0.0 | Token based |
Qwen: Qwen3 Embedding 4Bqwen/qwen3-embedding-4b |
Embeddings | $0.02 | $0.0 | Token based |
Qwen: Qwen3 Embedding 8Bqwen/qwen3-embedding-8b |
Embeddings | $0.01 | $0.0 | Token based |
Sentence Transformers: all-MiniLM-L12-v2sentence-transformers/all-minilm-l12-v2 |
Embeddings | $0.005 | $0.0 | Token based |
Sentence Transformers: all-MiniLM-L6-v2sentence-transformers/all-minilm-l6-v2 |
Embeddings | $0.005 | $0.0 | Token based |
Sentence Transformers: all-mpnet-base-v2sentence-transformers/all-mpnet-base-v2 |
Embeddings | $0.005 | $0.0 | Token based |
Sentence Transformers: multi-qa-mpnet-base-dot-v1sentence-transformers/multi-qa-mpnet-base-dot-v1 |
Embeddings | $0.005 | $0.0 | Token based |
Sentence Transformers: paraphrase-MiniLM-L6-v2sentence-transformers/paraphrase-minilm-l6-v2 |
Embeddings | $0.005 | $0.0 | Token based |
Thenlper: GTE-Basethenlper/gte-base |
Embeddings | $0.005 | $0.0 | Token based |
Thenlper: GTE-Largethenlper/gte-large |
Embeddings | $0.01 | $0.0 | Token based |
VoyageAI by MongoDB: voyage-4voyageai/voyage-4 |
Embeddings | $0.06 | $0.0 | Token based |
VoyageAI by MongoDB: voyage-4-largevoyageai/voyage-4-large |
Embeddings | $0.12 | $0.0 | Token based |
Reported cached input tokens are billed at 10% of the published input rate. Model availability can change. Read the quickstart or open the catalog API (JSON).
Bring your own model.
A dedicated-compute option, custom quoted
Deploy private weights on Actaserve-owned infrastructure, subject to model, licence and runtime qualification. Bring your checkpoint and serving requirements so we can scope model bring-up.
Reserve capacityReady for your first call?
Explore the docs