actaserve

Simple ways to
start building.

Use the model API at published rates. Reserve dedicated compute with a quote shaped around your workload.

Model API

Usage-based pricing

Connect through OpenAI- and Anthropic-compatible interfaces. The active model catalogue lists current rates and billing units.

  • Choose from the active model catalogue
  • Published pricing for each model
  • $5 in account credit after email verification

Model capabilities and pricing apply. API access does not reserve dedicated hardware or guarantee data residency.

Dedicated compute

Custom quote

Reservations open

Actaserve will own and operate the hardware. Choose root/bare-metal access or a managed API endpoint.

  • Hardware procurement, power and cooling
  • Model bring-up and ongoing operations
  • Capacity, location and access scoped per project

Dedicated capacity is reserved per contract, with the delivery date set in your quote. It is not ready-to-use inventory. Deployment terms.

Find your model. See your rate.

Published prices in USD. Token prices below are per one million tokens; other models use their listed billing unit. Check the model's supported options before sending a request.

533 models

Active public model prices in US dollars
ModelTypeInput / 1M tokensOutput / 1M tokensOther billing unit
Poolside: Laguna XS 2.1 (free)poolside/laguna-xs-2.1:free Language $0.0$0.0Token based
PrismML: Ternary Bonsai 2 27Bprism-ml/ternary-bonsai-2-27b Language $0.075$0.5Token based
Qwen2.5 72B Instructqwen/qwen-2.5-72b-instruct Language $0.36$0.4Token based
Qwen2.5 Coder 32B Instructqwen/qwen-2.5-coder-32b-instruct Language $0.66$1.0Token based
Qwen: Qwen2.5 7B Instructqwen/qwen-2.5-7b-instruct Language $0.1$0.2Token based
Qwen: Qwen2.5 VL 72B Instructqwen/qwen2.5-vl-72b-instruct Language $0.8$1.0Token based
Qwen: Qwen3 14Bqwen/qwen3-14b Language $0.12$0.24Token based
Qwen: Qwen3 235B A22B Instruct 2507qwen/qwen3-235b-a22b-2507 Language $0.09$0.55Token based
Qwen: Qwen3 235B A22B Thinking 2507qwen/qwen3-235b-a22b-thinking-2507 Language $0.45$3.5Token based
Qwen: Qwen3 30B A3Bqwen/qwen3-30b-a3b Language $0.12$0.5Token based
Qwen: Qwen3 30B A3B Instruct 2507qwen/qwen3-30b-a3b-instruct-2507 Language $0.1$0.3Token based
Qwen: Qwen3 32Bqwen/qwen3-32b Language $0.08$0.28Token based
Qwen: Qwen3.5-122B-A10Bqwen/qwen3.5-122b-a10b Language $0.26$2.08Token based
Qwen: Qwen3.5-27Bqwen/qwen3.5-27b Language $0.26$2.6Token based
Qwen: Qwen3.5-35B-A3Bqwen/qwen3.5-35b-a3b Language $0.08$0.75Token based
Qwen: Qwen3.5 397B A17Bqwen/qwen3.5-397b-a17b Language $0.55$3.5Token based
Qwen: Qwen3.5-9Bqwen/qwen3.5-9b Language $0.1$0.15Token based
Qwen: Qwen3.5-Flashqwen/qwen3.5-flash-02-23 Language $0.065$0.26Token based
Qwen: Qwen3.5 Plus 2026-02-15qwen/qwen3.5-plus-02-15 Language $0.26$1.56Token based
Qwen: Qwen3.5 Plus 2026-04-20qwen/qwen3.5-plus-20260420 Language $0.3$1.8Token based

Reported cached input tokens are billed at 10% of the published input rate. Model availability can change. Read the quickstart or open the catalog API (JSON).

Bring your own model.

A dedicated-compute option, custom quoted

Deploy private weights on Actaserve-owned infrastructure, subject to model, licence and runtime qualification. Bring your checkpoint and serving requirements so we can scope model bring-up.

Reserve capacity

Ready for your first call?

Explore the docs