CupAI logo

AI models and pricing

Access more than a hundred language, vision and audio models from OpenAI, Anthropic, Google and others through a single API key, paid in Toman and with no foreign bank account. Point the base URL of any OpenAI-compatible client at CupAI and you reach every model without changing a line of code, with all spending managed from one wallet.

Prices track each provider's own rate and are converted using the live exchange rate, and every request is billed from your wallet. Text models are billed on their input and output tokens; video models on the seconds they produce, and image models per generated image. Retired models are clearly marked and carry a suggested replacement, so moving to a newer version is straightforward.

0 models

No models published yet

The model catalogue will appear on this page shortly.

How model pricing works

Every AI model strikes a different balance between speed, answer quality and cost. Flagship models suit complex reasoning, long-document analysis and precise code generation, while lighter models cost considerably less for summarising, classification and short answers. The list above shows each model's rate broken down by input token, output token and cache, so you can estimate the cost of your workload before you start.

A token is the smallest unit of text a model processes; roughly speaking a thousand tokens is about 750 English words. The cost of a request is the sum of its input and output tokens, so shortening your prompt and capping the response length reduces cost directly. Some models have a second price tier for very long inputs, which reprices the entire request once it crosses a stated threshold.

Frequently asked questions

How is usage billed?
By the input and output tokens of each request, deducted from your wallet in Toman. Every model's rate is on this page.
Are the prices the same as the provider's?
Yes. The base price is the provider's own rate, converted to Toman at the live exchange rate.
Which model should I start with?
A fast mid-range model is enough for general work. If you need high accuracy or your input is very long, use one of the flagship models.
Do OpenAI-compatible clients work?
Yes. Just point the base URL at CupAI and use your own API key.