> For the complete documentation index, see [llms.txt](https://docs.warp.dev/llms.txt).
> Markdown versions of each page are available by appending .md to any URL.

# Usage and billing

Track Warp usage, understand inference, compute, and platform costs, and distinguish dollar billing from credit-based agreements.

Warp usage pays for model calls, hosted compute, and platform services. In the Warp app, open **Settings** > **Billing and usage** to check your included allowance, purchased balance, and reset date.

## Dollar usage and credits

On dollar-billed plans, your allowance and purchased usage are dollar amounts. Usage charges draw from that balance. Credit-based Enterprise agreements continue to use the rates and terms in your contract.

Usage has three charge components:

-   **Inference** - Model calls provided by Warp. If you use your own [API key](https://docs.warp.dev/agents/inference/bring-your-own-api-key/) or [inference endpoint](https://docs.warp.dev/agents/inference/custom-inference-endpoint/), your provider bills that inference separately.
-   **Compute** - The sandbox used by an agent on Warp-hosted compute. Local runs and self-hosted compute don’t incur this charge.
-   **Platform** - Agent time billed for cloud runs, plus local runs on Business and Enterprise that use customer-supplied inference. See [platform usage](https://docs.warp.dev/support-and-community/plans-and-billing/platform-credits/) for eligibility and rates.

For dollar-billed paid self-serve plans, Warp-provided inference and hosted compute use provider rates. Platform charges are separate. Free-plan [usage purchases](https://docs.warp.dev/support-and-community/plans-and-billing/add-on-credits/) include a purchase premium.

A credit is a usage unit, not a token or a fixed number of prompts. Dollar-billed plans can also display usage in credits; changing the display does not change what you’re charged.

## Tracking usage

-   **Account usage** - Check your balance and reset date in **Settings** > **Billing and usage**.
-   **Turn usage** - Expand the usage amount below an agent response. Where detailed usage is available, expand a model to see input, output, cache-read, and cache-write tokens, plus web searches and their costs.
-   **Conversation usage** - Open the conversation’s usage summary to see its total. **View account usage** opens Billing and usage.

On dollar-billed plans, choose **Credits** or **Dollars** in the “Usage display unit” dropdown in Billing and usage. Credit-billed plans remain credit-based. The CLI has its own [usage display](https://docs.warp.dev/agents/cli/models-and-usage/#usage-and-cost).

### Charges and estimates

Recorded dollar totals represent Warp usage charges, not the total price of your subscription or usage purchases. Older usage can have estimated costs, credit-only details, or no breakdown. An unavailable amount does not mean the run was free. Bills from your own inference provider are not fully represented in Warp’s totals.

## Included and purchased usage

On self-serve paid plans, included usage is allocated per seat and resets monthly, including on annual subscriptions. Unused included usage does not roll over. Check the allowance shown for your account rather than assuming it matches the amount advertised for a new subscription.

After included usage runs out, Warp draws from available [purchased usage](https://docs.warp.dev/support-and-community/plans-and-billing/add-on-credits/) on eligible plans. Purchased usage rolls over until its expiration and is shared across the team. Dedicated cloud allowances can be used before your general balance. Enterprise pool sizes and allocation rules follow your contract.

Dollar-billed Free plans don’t include bundled usage for the Warp Agent. Buy additional usage, [upgrade](https://www.warp.dev/pricing), or [bring your own inference](https://docs.warp.dev/agents/inference/bring-your-own-api-key/). Separate [platform charges](https://docs.warp.dev/support-and-community/plans-and-billing/platform-credits/) still apply to eligible runs.

## Other metered features

[Generate](https://docs.warp.dev/agents/local-agents/generate/) and [AI Autofill in Workflows](https://docs.warp.dev/knowledge-and-collaboration/warp-drive/workflows/#ai-autofill) also use inference. Regular shell commands and non-AI terminal features do not consume agent usage.

## How usage is calculated

Inference cost depends on the model’s rates and the work required to complete your request:

-   **Model choice** - Models have different input, output, and cache prices. A lower-cost model can reduce spending without reducing the number of tokens.
-   **Input and output** - Your prompt, conversation history, attached context, and generated response contribute to token usage.
-   **Caching** - Cached input can cost less than new input. Cache-read and cache-write tokens are reported separately.
-   **Task complexity** - A task can require several model calls, including calls made while using tools or summarizing a long conversation.
-   **Provider tools** - Features such as web search can add charges beyond token costs.

Two similar prompts can use different amounts. Use the reported breakdown to compare tasks rather than treating a prompt as a fixed-price unit. See [using tokens efficiently](https://docs.warp.dev/guides/configuration/how-to-use-tokens-efficiently-with-ai-coding-agents/) for ways to reduce usage.

## Compute usage

Cloud runs on Warp-hosted compute incur compute charges, whether started from the Warp app, an integration, the CLI, or the Warp Platform API. Compute cost depends on the resources and time used.

Local runs, CI jobs on your own runners, and [self-hosted workers](https://docs.warp.dev/factories/self-hosting/) do not incur Warp-hosted compute charges. Platform charges can still apply.

## Platform usage

Cloud runs, including factory runs and third-party harnesses, incur platform usage for billable agent time. Local runs on Business and Enterprise also incur platform charges when they use customer-supplied inference. Enterprise billing follows your contract.

See [platform usage](https://docs.warp.dev/support-and-community/plans-and-billing/platform-credits/) for billable time, exceptions, and the distinction between agent hours and run time.

## Cloud agent runs on team plans

Runs started with an agent API key use eligible shared team usage, not the team owner’s personal included allowance. Cloud agents don’t receive a separate per-seat allowance.

By default, user-created schedules use the creator’s eligible included and purchased usage. A schedule configured to run as a cloud agent uses shared team usage instead. Enterprise pools and charges follow your contract.

Review [auto-reload and purchase limits](https://docs.warp.dev/support-and-community/plans-and-billing/add-on-credits/#2-enable-auto-reload) before relying on unattended runs. If no usable balance remains, a run can fail with an [insufficient credits](https://docs.warp.dev/factories/api-and-sdk/troubleshooting/errors/insufficient-credits/) error.

## Related pages

-   [Plans, pricing, and refunds](https://docs.warp.dev/support-and-community/plans-and-billing/plans-pricing-refunds/) - Allowances, existing balances, and refund policies.
-   [Purchasing additional usage](https://docs.warp.dev/support-and-community/plans-and-billing/add-on-credits/) - One-time purchases, auto-reload, and team limits.
-   [Platform usage](https://docs.warp.dev/support-and-community/plans-and-billing/platform-credits/) - Agent-hour billing and eligibility.
