> ## Documentation Index
> Fetch the complete documentation index at: https://docs.capy.ai/llms.txt
> Use this file to discover all available pages before exploring further.

# Models & pricing

Capy bills for two things: model tokens and machine server time. Usage draws from one balance, in dollars, shared by your whole organization.

A subscription plan includes a set amount of usage each month at a discount. You can also add balance directly (minimum \$5) and enable [auto-reload](/admin/billing#auto-reload).

## Plans

Three plans, sized by the monthly credit grant. Annual billing is 12 months upfront at 20% off, with the same credits granted monthly.

| Plan           | Monthly       | Annual (per month) | Credits per month        |
| -------------- | ------------- | ------------------ | ------------------------ |
| **Capy Lite**  | \$20          | \$16               | \$20 of credits          |
| **Capy Pro**   | \$100         | \$80               | \$105 of credits (+5%)   |
| **Capy Max**   | \$200–\$1,000 | \$160–\$800        | 110% of the price (+10%) |
| **Enterprise** | Custom        | Custom             | Custom                   |

Capy Max comes in \$100 steps from \$200 to \$1,000. Every step grants 110% of its price in credits, so \$500/mo buys \$550 of credits.

**Try Capy Lite for \$1.** Your first 7 days cost \$1 with a verified card, then \$20/mo. One intro per card and per organization. Monthly Lite only.

Your credits pay for usage and the [Capy token rate](#capy-token-rate) alike. If you use them up mid-month, top up manually or with [auto-reload](/admin/billing#auto-reload).

## No seats

Every plan includes unlimited members. There is no per-seat price and no charge for inviting anyone. Plans size usage, nothing else.

## Models

Picking a model entry picks both the model and how usage gets billed. `openai/gpt-5.6-sol` bills your Capy balance at the rates below. `codex/gpt-5.6-sol` runs the same model on your ChatGPT subscription.

Prices are per 1M tokens. Models marked BYOK can also run on your own provider API key from [Settings → Models](https://capy.ai/settings/models). The token rate column is the flat [Capy Token Rate](#capy-token-rate) per million tokens.

| Model                      | ID                                | Context | Input (per 1M)                                                    | Output (per 1M)                      | Token rate (per 1M) | BYOK |
| -------------------------- | --------------------------------- | ------- | ----------------------------------------------------------------- | ------------------------------------ | ------------------- | ---- |
| **GPT-5.6 Sol** (default)  | `openai/gpt-5.6-sol`              | 1.05M   | ≤272K: \$5.00 (cached \$0.50)<br />>272K: \$10.00 (cached \$1.00) | ≤272K: \$30.00<br />>272K: \$45.00   | \$0.20              | ✓    |
| **GPT-5.6 Terra**          | `openai/gpt-5.6-terra`            | 1.05M   | ≤272K: \$2.50 (cached \$0.25)<br />>272K: \$5.00 (cached \$0.50)  | ≤272K: \$15.00<br />>272K: \$22.50   | \$0.20              | ✓    |
| **GPT-5.6 Luna**           | `openai/gpt-5.6-luna`             | 1.05M   | ≤272K: \$1.00 (cached \$0.10)<br />>272K: \$2.00 (cached \$0.20)  | ≤272K: \$6.00<br />>272K: \$9.00     | \$0.02              | ✓    |
| **GPT-5.5**                | `openai/gpt-5.5`                  | 1.05M   | ≤272K: \$5.00 (cached \$0.50)<br />>272K: \$10.00 (cached \$1.00) | ≤272K: \$30.00<br />>272K: \$45.00   | \$0.20              | ✓    |
| **GPT-5.5 Pro**            | `openai/gpt-5.5-pro`              | 1.05M   | ≤272K: \$30.00<br />>272K: \$60.00                                | ≤272K: \$180.00<br />>272K: \$270.00 | \$0.20              | ✓    |
| **GPT-5.4 Mini**           | `openai/gpt-5.4-mini`             | 400K    | \$0.75 (cached \$0.075)                                           | \$4.50                               | \$0.02              | ✓    |
| **GPT-5.4**                | `openai/gpt-5.4`                  | 1.05M   | ≤272K: \$2.50 (cached \$0.25)<br />>272K: \$5.00 (cached \$0.50)  | ≤272K: \$15.00<br />>272K: \$22.50   | \$0.20              | ✓    |
| **GPT-5.3 Codex**          | `openai/gpt-5.3-codex`            | 400K    | \$1.75 (cached \$0.175)                                           | \$14.00                              | \$0.02              | ✓    |
| **Claude Fable 5**         | `anthropic/claude-fable-5`        | 1M      | \$10.00 (cached \$1.00)                                           | \$50.00                              | \$0.20              | ✓    |
| **Claude Opus 5**          | `anthropic/claude-opus-5`         | 1M      | \$5.00 (cached \$0.50)                                            | \$25.00                              | \$0.20              | ✓    |
| **Claude Sonnet 5**        | `anthropic/claude-sonnet-5`       | 1M      | \$2.00 (cached \$0.20)                                            | \$10.00                              | \$0.02              | ✓    |
| **Claude Haiku 4.5**       | `anthropic/claude-haiku-4-5`      | 200K    | \$1.00 (cached \$0.10)                                            | \$5.00                               | \$0.02              | ✓    |
| **Claude Opus 4.8**        | `anthropic/claude-opus-4-8`       | 1M      | \$5.00 (cached \$0.50)                                            | \$25.00                              | \$0.20              | ✓    |
| **Claude Opus 4.7**        | `anthropic/claude-opus-4-7`       | 1M      | \$5.00 (cached \$0.50)                                            | \$25.00                              | \$0.20              | ✓    |
| **Claude Sonnet 4.6**      | `anthropic/claude-sonnet-4-6`     | 1M      | \$3.00 (cached \$0.30)                                            | \$15.00                              | \$0.20              | ✓    |
| **Claude Opus 4.6**        | `anthropic/claude-opus-4-6`       | 1M      | \$5.00 (cached \$0.50)                                            | \$25.00                              | \$0.20              | ✓    |
| **Claude Opus 4.5**        | `anthropic/claude-opus-4-5`       | 200K    | \$5.00 (cached \$0.50)                                            | \$25.00                              | \$0.20              | ✓    |
| **Grok 4.6**               | `xai/grok-4.6`                    | 500K    | \<200K: \$2.00 (cached \$0.50)<br />≥200K: \$4.00 (cached \$1.00) | \<200K: \$6.00<br />≥200K: \$12.00   | \$0.02              | ✓    |
| **Grok 4.5**               | `xai/grok-4.5`                    | 500K    | ≤200K: \$2.00 (cached \$0.30)<br />>200K: \$4.00 (cached \$0.60)  | ≤200K: \$6.00<br />>200K: \$12.00    | \$0.02              | ✓    |
| **Grok 4.3**               | `xai/grok-4.3`                    | 1M      | ≤200K: \$1.25 (cached \$0.20)<br />>200K: \$2.50 (cached \$0.40)  | ≤200K: \$2.50<br />>200K: \$5.00     | \$0.02              | ✓    |
| **Kimi K3**                | `moonshotai/kimi-k3`              | 1.05M   | \$3.00 (cached \$0.30)                                            | \$15.00                              | \$0.20              |      |
| **Kimi K2.7 Code**         | `moonshotai/kimi-k2.7-code`       | 262K    | \$0.95 (cached \$0.19)                                            | \$4.00                               | \$0.02              |      |
| **Kimi K2.6**              | `moonshotai/kimi-k2.6`            | 262K    | \$0.95 (cached \$0.16)                                            | \$4.00                               | \$0.02              |      |
| **GLM-5.2**                | `zai/glm-5.2`                     | 1.05M   | \$1.40 (cached \$0.26)                                            | \$4.40                               | \$0.02              |      |
| **GLM-5.1**                | `zai/glm-5.1`                     | 203K    | \$1.40 (cached \$0.26)                                            | \$4.40                               | \$0.02              |      |
| **GLM-5V-Turbo**           | `zai/glm-5v-turbo`                | 203K    | \$1.20 (cached \$0.24)                                            | \$4.00                               | \$0.02              |      |
| **GLM-5-Turbo**            | `zai/glm-5-turbo`                 | 203K    | \$1.20 (cached \$0.24)                                            | \$4.00                               | \$0.02              |      |
| **DeepSeek V4 Pro**        | `deepseek/deepseek-v4-pro`        | 1.05M   | \$1.74 (cached \$0.145)                                           | \$3.48                               | \$0.02              |      |
| **DeepSeek V4 Flash 0731** | `deepseek/deepseek-v4-flash-0731` | 1.05M   | \$0.14 (cached \$0.028)                                           | \$0.28                               | \$0.02              |      |
| **Gemini 3.1 Pro Preview** | `google/gemini-3.1-pro-preview`   | 1.05M   | ≤200K: \$2.00 (cached \$0.20)<br />>200K: \$4.00 (cached \$0.40)  | ≤200K: \$12.00<br />>200K: \$18.00   | \$0.02              | ✓    |
| **Gemini 3 Flash Preview** | `google/gemini-3-flash-preview`   | 1.05M   | \$0.50 (cached \$0.05)                                            | \$3.00                               | \$0.02              | ✓    |
| **Qwen3.8 Max**            | `qwen/qwen3.8-max`                | 1M      | \$2.00 (cached \$0.25)                                            | \$6.00                               | \$0.02              |      |

## OAuth subscriptions

These entries run through a connected OAuth subscription (Codex, Copilot, SuperGrok) or an organization account (Azure). Tokens bill to that subscription instead of your Capy balance. Usage on external subscriptions will incur an additional [Capy Token Rate](#capy-token-rate).

| Model                      | ID                               | Runs through         | Context | Token rate (per 1M) |
| -------------------------- | -------------------------------- | -------------------- | ------- | ------------------- |
| **GPT-5.6 Sol**            | `codex/gpt-5.6-sol`              | Codex                | 400K    | \$0.20              |
| **GPT-5.6 Terra**          | `codex/gpt-5.6-terra`            | Codex                | 400K    | \$0.20              |
| **GPT-5.6 Luna**           | `codex/gpt-5.6-luna`             | Codex                | 400K    | \$0.02              |
| **GPT-5.5**                | `codex/gpt-5.5`                  | Codex                | 400K    | \$0.20              |
| **GPT-5.4 Mini**           | `codex/gpt-5.4-mini`             | Codex                | 400K    | \$0.02              |
| **GPT-5.4**                | `codex/gpt-5.4`                  | Codex                | 400K    | \$0.20              |
| **GPT-5.3 Codex Spark**    | `codex/gpt-5.3-codex-spark`      | Codex                | 128K    | \$0.20              |
| **GPT-5.6 Sol**            | `copilot/gpt-5.6-sol`            | Copilot              | 1.05M   | \$0.20              |
| **GPT-5.6 Terra**          | `copilot/gpt-5.6-terra`          | Copilot              | 1.05M   | \$0.20              |
| **GPT-5.6 Luna**           | `copilot/gpt-5.6-luna`           | Copilot              | 1.05M   | \$0.02              |
| **Claude Opus 4.8**        | `copilot/claude-opus-4-8`        | Copilot              | 200K    | \$0.20              |
| **GPT-5.5**                | `copilot/gpt-5.5`                | Copilot              | 1.05M   | \$0.20              |
| **Claude Opus 4.7**        | `copilot/claude-opus-4-7`        | Copilot              | 200K    | \$0.20              |
| **GPT-5.4 Mini**           | `copilot/gpt-5.4-mini`           | Copilot              | 400K    | \$0.02              |
| **GPT-5.4**                | `copilot/gpt-5.4`                | Copilot              | 1.05M   | \$0.20              |
| **Gemini 3.1 Pro Preview** | `copilot/gemini-3.1-pro-preview` | Copilot              | 1M      | \$0.02              |
| **Claude Sonnet 4.6**      | `copilot/claude-sonnet-4-6`      | Copilot              | 200K    | \$0.20              |
| **GPT-5.3 Codex**          | `copilot/gpt-5.3-codex`          | Copilot              | 400K    | \$0.02              |
| **Claude Opus 4.6**        | `copilot/claude-opus-4-6`        | Copilot              | 200K    | \$0.20              |
| **Gemini 3 Flash Preview** | `copilot/gemini-3-flash-preview` | Copilot              | 128K    | \$0.02              |
| **Claude Opus 4.5**        | `copilot/claude-opus-4-5`        | Copilot              | 200K    | \$0.20              |
| **Claude Haiku 4.5**       | `copilot/claude-haiku-4-5`       | Copilot              | 200K    | \$0.02              |
| **Grok 4.5**               | `supergrok/grok-4.5`             | SuperGrok            | 500K    | \$0.02              |
| **Composer 2.5 Fast**      | `supergrok/composer-2.5-fast`    | SuperGrok            | 200K    | \$0.20              |
| **GPT-5.6 Sol**            | `azure/gpt-5.6-sol`              | Azure (organization) | 1.05M   | \$0.20              |
| **Claude Fable 5**         | `azure/claude-fable-5`           | Azure (organization) | 1M      | \$0.20              |

## Fast mode

Models with a Fast toggle run through the provider's priority tier at a higher token rate. When no fast host is available, the request serves at the standard tier and bills at standard rates. Fast rates per 1M tokens:

| Model               | Fast input (per 1M)                                               | Fast output (per 1M)                |
| ------------------- | ----------------------------------------------------------------- | ----------------------------------- |
| **GPT-5.6 Sol**     | \$10.00 (cached \$1.00)                                           | \$60.00                             |
| **GPT-5.6 Terra**   | \$5.00 (cached \$0.50)                                            | \$30.00                             |
| **GPT-5.6 Luna**    | \$2.00 (cached \$0.20)                                            | \$12.00                             |
| **GPT-5.5**         | \$12.50 (cached \$1.25)                                           | \$75.00                             |
| **GPT-5.4 Mini**    | \$1.50 (cached \$0.15)                                            | \$9.00                              |
| **GPT-5.4**         | \$5.00 (cached \$0.50)                                            | \$30.00                             |
| **GPT-5.3 Codex**   | \$3.50 (cached \$0.35)                                            | \$28.00                             |
| **Claude Opus 5**   | \$10.00 (cached \$1.00)                                           | \$50.00                             |
| **Claude Opus 4.8** | \$10.00 (cached \$1.00)                                           | \$50.00                             |
| **Grok 4.6**        | \<200K: \$4.00 (cached \$1.00)<br />≥200K: \$8.00 (cached \$2.00) | \<200K: \$12.00<br />≥200K: \$24.00 |

## Reasoning effort

Most models expose a reasoning effort, from `none` up to `max`. Each entry supports its own subset, shown in the picker. An unsupported effort fails before the request is sent. Capy never quietly approximates it with a different setting.

## Capy Token Rate

On-demand model requests include a Capy Token Rate of \$0.20 to \$0.02 per million tokens. This rate applies on top of model API pricing for on-demand usage, subscription usage, and BYOK usage. The [models table](#models) lists each model's rate.

## Bring your own key (BYOK)

BYOK can be configured for enterprise organizations, replacing standard Capy model routing with your own configured key and API endpoint. Activate it in [Settings → Models](https://capy.ai/settings/models) with your provider API key. Your key pays the provider directly, and your balance is charged only the [Capy Token Rate](#capy-token-rate).

## Per-task models

Threads and tasks pick models independently. The agent can run a task on a cheaper model for mechanical work, or a stronger one for a hard subsystem. Tasks follow parent thread models by default.

## When your balance runs out

A run that can't pass the balance check stops with a visible error, recorded in the thread. Add balance manually or with [auto-reload](/admin/billing#auto-reload), then retry or resend. The thread picks up from its recorded history.

## Machine pricing

Each thread runs on its own VM machine, billed hourly by size:

| Size                 | vCPU | RAM    | Disk   | Cost/hour |
| -------------------- | ---- | ------ | ------ | --------- |
| Small                | 1    | 4 GB   | 64 GB  | \$0.10    |
| Medium               | 2    | 8 GB   | 64 GB  | \$0.20    |
| Large                | 4    | 16 GB  | 64 GB  | \$0.40    |
| Ultra                | 8    | 32 GB  | 64 GB  | \$0.80    |
| Hyper                | 16   | 64 GB  | 128 GB | \$1.60    |
| Big Guy (enterprise) | 16   | 128 GB | 256 GB | \$3.20    |

Large is the default. You pay while a machine is awake, plus a two-minute wind-down when it goes to sleep. A sleeping machine costs nothing, but waking it to extend its uptime for ports, desktop, file viewer will incur a charge. See [machines](/machines).

## Enterprise

Need more than Capy Max's \$1,000/month in credits, custom pricing, or invoiced billing with net terms? We set up a custom plan. Everything else works the same as on standard plans.

[Contact us →](https://cal.com/team/capy/enterprise)

## FAQ

<AccordionGroup>
  <Accordion title="How does the $1 intro work?">
    Subscribe to monthly Capy Lite with a verified card and your first 7 days cost \$1, then \$20/mo unless you cancel. One intro per card and per organization, ever.
  </Accordion>

  <Accordion title="Do members cost anything?">
    No. Every plan includes unlimited members with no per-seat price. Inviting someone never costs anything; only usage does.
  </Accordion>

  <Accordion title="What is the Capy Token Rate?">
    A flat rate per million tokens on Capy-hosted usage, BYOK, and subscription routes. Your balance pays it like any other charge. Per-model rates are in the [models table](#models).
  </Accordion>

  <Accordion title="What happens when my balance runs out?">
    Active runs stop with a visible error in the thread. Add balance or let auto-reload top up, then retry the thread. Nothing resumes on its own.
  </Accordion>

  <Accordion title="Can I switch plans?">
    Yes, from the billing page. Upgrades apply immediately with a prorated charge. Downgrades take effect at the end of the billing cycle.
  </Accordion>

  <Accordion title="How do annual plans work?">
    Annual plans bill upfront for 12 months at a 20% discount. Credits are still granted monthly.
  </Accordion>

  <Accordion title="Do I need a subscription at all?">
    No, you can add balance directly (minimum \$5). A plan grants credits monthly at a discount.
  </Accordion>
</AccordionGroup>
