Billing formula
Total Cost = (input_tokens / 1,000,000 * input_price) + (output_tokens / 1,000,000 * output_price)
Cache read and cache write prices apply only when the upstream route reports those billing buckets.
OpenAI
Anthropic
Google Gemini
xAI
DeepSeek
Alibaba
Moonshot
MiniMax
Zhipu GLM
ByteDance
Xiaomi
StepFun
HappyHorse
Other
Important notes
- Use aisa.one/models for the latest live availability and pricing before production changes.
- Token models are billed on input and output usage. Cache read and cache write prices apply only when the upstream route reports those billing buckets.
- Image models are billed per request or per generated image.
- Video models use two different meters: Wan and HappyHorse routes bill per second of output with a separate rate per resolution, while the Dreamina Seedance routes bill per 1M tokens with a separate rate per resolution.
- Embedding models are billed on input tokens only.
gpt-image-2publishes both token prices and per-request / per-image tiers; the tier that applies depends on the route you call.- The final billed amount for each call is visible in the AIsa Usage Logs page.