CupAI logo

← Back to all models

Moonshot AI logo

Kimi K2.7 Code

⁦Moonshot AI⁩ · ⁦kimi-k2.7-code⁩

Moonshot AI's open-weight coding model for agentic programming; text and image input, 256K context, always-on thinking.

⁦262K⁩ tokensText, Image → TextInput/output type

Starting price

241,395

per million tokens

Billing

Pay as you go

Service status

Active

Released

2026

About Kimi K2.7 Code

Moonshot AI's open-weight coding model for agentic programming; text and image input, 256K context, always-on thinking.

Kimi K2.7 Code pricing

Input

241,395

Output

1,016,400

Cache (read)

48,279

Cache (write)

241,395

Model id

kimi-k2.7-code

Provider

Moonshot AI

Context window

262K

Max output tokens

262,144

Inputs

Text, Image

Outputs

Text

Billing basis

Input and output tokens

Status

Active

Released

2026

Share model

https://cupai.ir/en/models/kimi-k2.7-code

Sample code and API for Kimi K2.7 Code

Point the base URL at CupAI and use your own API key; the rest of the request stays exactly as it is in any compatible client.

curl 'https://api.cupai.ir/v1/chat/completions' \
  -H 'Content-Type: application/json' \
  -H 'Authorization: Bearer $CUPAI_API_KEY' \
  -d '{
    "model": "kimi-k2.7-code",
    "messages": [
      {
        "role": "user",
        "content": "Hello!"
      }
    ]
  }'

Replace CUPAI_API_KEY with your own key.

Create an API key

Capabilities

Capabilities supported by this model's API.

text-to-text

image-to-text

Frequently asked questions about Kimi K2.7 Code

How is usage of Kimi K2.7 Code billed?
By the input and output tokens of each request. The rate for each part is in the pricing table on this page, and the amount is deducted from your wallet.
How do I connect to Kimi K2.7 Code?
Point the base URL at CupAI, put your API key in the request header, and send the model id in the request body. Ready-made samples are on this page.
How do payments work?
Payments are in Toman and need no foreign credit card. Credit you buy can be spent on any available model.

Other models

How model pricing works

Every AI model strikes a different balance between speed, answer quality and cost. Flagship models suit complex reasoning, long-document analysis and precise code generation, while lighter models cost considerably less for summarising, classification and short answers. The list above shows each model's rate broken down by input token, output token and cache, so you can estimate the cost of your workload before you start.

A token is the smallest unit of text a model processes; roughly speaking a thousand tokens is about 750 English words. The cost of a request is the sum of its input and output tokens, so shortening your prompt and capping the response length reduces cost directly. Some models have a second price tier for very long inputs, which reprices the entire request once it crosses a stated threshold.