# Models and prices

**Beta.** Prices on this page are the ones the API charges.

## Prices

Each price is per million tokens, in US dollars: the provider's list price, plus our 5%, equals what you pay. You're charged in sats from your balance.

| Model | Provider | Input | Cached input | Output |
| --- | --- | --- | --- | --- |
| `google/gemini-2.5-flash-lite` | Google Vertex AI | $0.1 + $0.005 = **$0.105** | $0.01 + $0.0005 = **$0.0105** | $0.4 + $0.02 = **$0.42** |
| `google/gemini-3.8-flash` | OpenRouter | $0.75 + $0.0375 = **$0.7875** | $0.075 + $0.00375 = **$0.07875** | $3.75 + $0.1875 = **$3.9375** |
| `google/gemini-3.8-flash` | Vercel AI Gateway | $0.75 + $0.0375 = **$0.7875** | $0.075 + $0.00375 = **$0.07875** | $3.75 + $0.1875 = **$3.9375** |
| `google/gemini-3.8-flash` | Google Vertex AI | $0.75 + $0.0375 = **$0.7875** | $0.075 + $0.00375 = **$0.07875** | $3.75 + $0.1875 = **$3.9375** |
| `openai/gpt-5.6-luna` | OpenAgents (Pro) | $0.2 + $0.01 = **$0.21** | $0.02 + $0.001 = **$0.021** | $1.2 + $0.06 = **$1.26** |
| `openai/gpt-5.6-sol` | OpenAgents (Pro) | $4 + $0.2 = **$4.2** | $0.4 + $0.02 = **$0.42** | $20 + $1 = **$21** |
| `openai/gpt-5.6-terra` | OpenAgents (Pro) | $2 + $0.1 = **$2.1** | $0.2 + $0.01 = **$0.21** | $12 + $0.6 = **$12.6** |
| `stealth/space-bunny-alpha` | OpenRouter | $0 + $0 = **$0** | $0 + $0 = **$0** | $0 + $0 = **$0** |
| `zai/glm-5.3-flash` | OpenRouter | $0.15 + $0.0075 = **$0.1575** | $0.03 + $0.0015 = **$0.0315** | $0.5 + $0.025 = **$0.525** |
| `zai/glm-5.3-flash` | Vercel AI Gateway | $0.15 + $0.0075 = **$0.1575** | $0.03 + $0.0015 = **$0.0315** | $0.5 + $0.025 = **$0.525** |
| `zai/glm-5.3-flash` | Z.ai | $0.15 + $0.0075 = **$0.1575** | $0.03 + $0.0015 = **$0.0315** | $0.5 + $0.025 = **$0.525** |


The same card is JSON at `GET /v1/rates`, open without a key. Each model in
`GET /v1/models` carries its rows too.

A few rules:

- **The list price, whoever answers.** When a model has more than one
  provider, you pay that provider's row. Our own deals with providers never
  change the price you see.
- **Promotions are their own rows**, labeled, beside the list price.
- **Your own key costs nothing extra.** See
  [Bring your own key](/docs/api/bring-your-own-key).

## What each model can do

| Model | Context | Most output | Tools | Images | JSON schema |
| --- | --- | --- | --- | --- | --- |
| `google/gemini-3.8-flash` | 1,048,576 | 65,536 | Yes | Yes | Yes |
| `google/gemini-2.5-flash-lite` | 1,048,576 | 65,536 | Yes | Yes | Yes |
| `zai/glm-5.3-flash` | 1,000,000 | 128,000 | Yes | No | No |
| `openai/gpt-5.6-sol` | 400,000 | 128,000 | No | No | No |
| `openai/gpt-5.6-terra` | 400,000 | 128,000 | No | No | No |
| `openai/gpt-5.6-luna` | 400,000 | 128,000 | No | No | No |
| `stealth/space-bunny-alpha` | 256,000 | 64,000 | Yes | No | Yes |


## Tokens served

**0 tokens served** (input and output).

| Caller | Free calls | Paid calls |
| --- | ---: | ---: |
| Our services | 0 | 0 |
| Outside callers | 0 | 0 |

Paid calls include **0 tokens on callers' own keys**.

[Read the same counts as JSON](/api/v1/usage/tokens-served). Only answers with reported token counts are included.

## Model names

Model names are `publisher/model`, such as `zai/glm-5.3-flash`. Names that
start with `openagents/` pick a model for you by task:

| Name | For |
| --- | --- |
| `openagents/auto` | We read the request and pick the task below |
| `openagents/classify` | Labels, extraction, a short yes or no |
| `openagents/fast` | Short replies, openers, summaries |
| `openagents/chat` | General conversation |
| `openagents/code` | Coding and tool use |
| `openagents/long` | Inputs over 200,000 tokens |
| `openagents/reason` | Hard multi-step problems |

[Which model to use](/docs/api/decisions) helps you choose.
