Skip to main content
This page lists current public AIsa Model Gateway prices for all 108 models in the gateway, refreshed from the live pricing feed on August 5, 2026. For model capabilities, context windows, and endpoint mappings, see the supported model catalog. All token prices are in USD per 1 million tokens. Per-request models are billed per generated asset or call. Workspace-level pricing rules can change the final amount shown in Usage Logs.

Billing formula

Total Cost = (input_tokens / 1,000,000 * input_price) + (output_tokens / 1,000,000 * output_price) Cache read and cache write prices apply only when the upstream route reports those billing buckets.

OpenAI

Anthropic

Google Gemini

xAI

DeepSeek

Alibaba

Moonshot

MiniMax

Zhipu GLM

ByteDance

Xiaomi

StepFun

HappyHorse

Other

Important notes

  • Use aisa.one/models for the latest live availability and pricing before production changes.
  • Token models are billed on input and output usage. Cache read and cache write prices apply only when the upstream route reports those billing buckets.
  • Image models are billed per request or per generated image.
  • Video models use two different meters: Wan and HappyHorse routes bill per second of output with a separate rate per resolution, while the Dreamina Seedance routes bill per 1M tokens with a separate rate per resolution.
  • Embedding models are billed on input tokens only.
  • gpt-image-2 publishes both token prices and per-request / per-image tiers; the tier that applies depends on the route you call.
  • The final billed amount for each call is visible in the AIsa Usage Logs page.