[Docs index](/docs.md) / [Workspace Billing](/docs/workspace-billing/overview.md) / Viewing Usage and Costs

---

# Viewing Usage and Costs

Monitor your workspace's AI and tool usage, costs, and execution history from the Usage page.

## Before you begin

- You must be a workspace administrator to view workspace-level usage data.

## Steps

### 1. Open the Usage page

Navigate to the **Usage** page from the main navigation. The page shows a summary of your workspace's spending and activity.

### 2. Review the usage chart

The stacked bar chart displays monthly costs broken down by:

- **LLM Cost** (cyan bars) -- token costs for AI model usage
- **Tool Cost** (orange bars) -- costs for tool executions

Hover over any bar to see a detailed tooltip with the month's total cost, LLM cost, tool cost, execution count, and success/failure breakdown.

Above the chart, three summary metrics are displayed:
- **Total spend** across the displayed period
- **Average monthly cost**
- **Total executions**

### 3. Check the stats cards

Four cards show key metrics at a glance:

- **Current Month** -- spending so far this month, with a trend indicator comparing to last month
- **Last Month** -- total spending from the previous month
- **Success Rate** -- percentage of successful tool executions
- **Recent Failures** -- number of failed executions in the last 24 hours

### 4. Review recent executions

The **Recent Executions** section lists the most recent tool runs. Each entry shows:

- Status (green for succeeded, red for failed, orange for in progress)
- Tool name
- User who triggered it
- Time elapsed since execution
- Duration
- Tool pack name

Click **See more** to view additional executions if available.

### 5. View pricing details

Click the **Pricing** button (blue info icon) in the page header to open the pricing dialog. This shows:

- **LLM Pricing** -- input and output token costs per million tokens for each model
- **Tool Pricing** -- per-call rates for tool executions

## Model pricing reference

The table below lists all supported models and their token costs. These are the base costs before any workspace billing margin is applied.

### OpenAI

| Model | Input per 1M tokens | Output per 1M tokens | Source |
|-------|---------------------:|----------------------:|--------|
| gpt-5.2-codex | $1.75 | $14.00 | [OpenAI Pricing](https://openai.com/api/pricing/) |
| gpt-5 | $1.25 | $10.00 | [OpenAI Pricing](https://openai.com/api/pricing/) |
| gpt-5-mini | $0.25 | $2.00 | [OpenAI Pricing](https://openai.com/api/pricing/) |
| gpt-5-nano | $0.05 | $0.40 | [OpenAI Pricing](https://openai.com/api/pricing/) |
| 4.1 | $2.00 | $8.00 | [OpenAI Pricing](https://openai.com/api/pricing/) |
| 4o | $2.50 | $10.00 | [OpenAI Pricing](https://openai.com/api/pricing/) |
| 4o-mini | $0.15 | $0.60 | [OpenAI Pricing](https://openai.com/api/pricing/) |
| o4-mini | $1.10 | $4.40 | [OpenAI Pricing](https://openai.com/api/pricing/) |
| o3-mini | $1.10 | $4.40 | [OpenAI Pricing](https://openai.com/api/pricing/) |

### Anthropic

| Model | Input per 1M tokens | Output per 1M tokens | Source |
|-------|---------------------:|----------------------:|--------|
| claude-fable-5 | $10.00 | $50.00 | [Anthropic Pricing](https://platform.claude.com/docs/en/about-claude/pricing) |
| claude-opus-4-6 | $5.00 | $25.00 | [Anthropic Pricing](https://platform.claude.com/docs/en/about-claude/pricing) |
| claude-sonnet-5 | $3.00 | $15.00 | [Anthropic Pricing](https://platform.claude.com/docs/en/about-claude/pricing) |
| claude-sonnet-4-6 | $3.00 | $15.00 | [Anthropic Pricing](https://platform.claude.com/docs/en/about-claude/pricing) |
| claude-opus-4-5 | $5.00 | $25.00 | [Anthropic Pricing](https://platform.claude.com/docs/en/about-claude/pricing) |
| claude-sonnet-4-5 | $3.00 | $15.00 | [Anthropic Pricing](https://platform.claude.com/docs/en/about-claude/pricing) |
| claude-haiku-4-5 | $1.00 | $5.00 | [Anthropic Pricing](https://platform.claude.com/docs/en/about-claude/pricing) |
| claude-opus-4.1 | $15.00 | $75.00 | [Anthropic Pricing](https://platform.claude.com/docs/en/about-claude/pricing) |
| claude-opus-4 | $15.00 | $75.00 | [Anthropic Pricing](https://platform.claude.com/docs/en/about-claude/pricing) |
| claude-sonnet-4 | $3.00 | $15.00 | [Anthropic Pricing](https://platform.claude.com/docs/en/about-claude/pricing) |
| claude-3-7-sonnet | $3.00 | $15.00 | [Anthropic Pricing](https://platform.claude.com/docs/en/about-claude/pricing) |

### xAI

| Model | Input per 1M tokens | Output per 1M tokens | Source |
|-------|---------------------:|----------------------:|--------|
| grok-3 | $3.00 | $15.00 | [xAI Pricing](https://docs.x.ai/developers/models) |
| grok-3-mini | $0.30 | $0.50 | [xAI Pricing](https://docs.x.ai/developers/models) |
| grok-4-1-fast-reasoning | $0.20 | $0.50 | [xAI Pricing](https://docs.x.ai/developers/models) |
| grok-4-1-fast-non-reasoning | $0.20 | $0.50 | [xAI Pricing](https://docs.x.ai/developers/models) |
| grok-code-fast-1 | $0.20 | $1.50 | [xAI Pricing](https://docs.x.ai/developers/models) |

### Perplexity

| Model | Input per 1M tokens | Output per 1M tokens | Source |
|-------|---------------------:|----------------------:|--------|
| sonar | $1.00 | $1.00 | [Perplexity Pricing](https://docs.perplexity.ai/docs/getting-started/pricing) |

### DeepInfra-hosted Models

| Model | Input per 1M tokens | Output per 1M tokens | Source |
|-------|---------------------:|----------------------:|--------|
| deepseek-ai/DeepSeek-V4-Flash | $0.10 | $0.20 | [DeepInfra Pricing](https://deepinfra.com/deepseek-ai/DeepSeek-V4-Flash) |
| deepseek-ai/DeepSeek-V4.1-Flash | $0.14 | $0.42 | [DeepInfra Pricing](https://deepinfra.com/deepseek-ai/DeepSeek-V4.1-Flash) |
| XiaomiMiMo/MiMo-V2.6-Flash | $0.14 | $0.28 | [DeepInfra Pricing](https://deepinfra.com/XiaomiMiMo/MiMo-V2.6-Flash) |
| XiaomiMiMo/MiMo-V2.6-Pro | $0.43 | $0.87 | [DeepInfra Pricing](https://deepinfra.com/XiaomiMiMo/MiMo-V2.6-Pro) |
| nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B | $0.50 | $2.50 | [DeepInfra Pricing](https://deepinfra.com/nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B) |
| zai-org/GLM-4.7-Flash | $0.06 | $0.40 | [DeepInfra Pricing](https://deepinfra.com/zai-org/GLM-4.7-Flash) |
| zai-org/GLM-5.1 | $1.05 | $3.50 | [DeepInfra Pricing](https://deepinfra.com/zai-org/GLM-5.1) |
| Qwen/Qwen3.7-Max | $2.50 | $7.50 | [DeepInfra Pricing](https://deepinfra.com/Qwen/Qwen3.7-Max) |
| Qwen/Qwen3.8-2.4T-A95B | $2.00 | $6.00 | [DeepInfra Pricing](https://deepinfra.com/Qwen/Qwen3.8-2.4T-A95B) |
| moonshotai/Kimi-K2.6 (FP4) | $0.75 | $3.50 | [DeepInfra Pricing](https://deepinfra.com/moonshotai/Kimi-K2.6) |
| zai-org/GLM-5.3 (FP4) | $0.563 | $2.50 | [DeepInfra Pricing](https://deepinfra.com/zai-org/GLM-5.3) |

### Cerebras

| Model | Input per 1M tokens | Output per 1M tokens | Source |
|-------|---------------------:|----------------------:|--------|
| zai-glm-4.7 | $2.25 | $2.75 | [Cerebras Pricing](https://www.cerebras.ai/pricing) |

### Self-hosted (SGLang)

Self-hosted models run on your own inference server, so there is no per-token
vendor charge and token costs are recorded as zero. Tool executions are still
billed at the usual rate.

| Model | Input per 1M tokens | Output per 1M tokens | Source |
|-------|---------------------:|----------------------:|--------|
| gemma-4-26b | $0.00 | $0.00 | Self-hosted |
| qwen3.8-27b | $0.00 | $0.00 | Self-hosted |

### Tool execution pricing

Tool executions are billed at a flat rate per call. The default global rate is $170.00 per million calls ($0.00017 per call). Individual tools or tool packs may have custom rates.

## Viewing per-chat token usage

Token counts appear alongside individual chat conversations. The count shows the total tokens used in that chat session, formatted as a compact number (for example, "1.2k tokens" or "2.5M tokens").

See [Troubleshooting](troubleshooting.md) if usage data is not loading.

---

## Navigation

### In this section: Workspace Billing

- [Workspace Billing](/docs/workspace-billing/overview.md)
- [Adding Credits](/docs/workspace-billing/adding-credits.md)
- [Setting Up a Payment Method](/docs/workspace-billing/setting-up-a-payment-method.md)
- [Configuring Automatic Top-Ups](/docs/workspace-billing/configuring-automatic-top-ups.md)
- **Viewing Usage and Costs** (current)
- [Troubleshooting](/docs/workspace-billing/troubleshooting.md)

### Other sections

- [Tool Creation](/docs/tool-creation/overview.md)
- [Subagents](/docs/subagents/overview.md)
- [Agent Skills](/docs/agent-skills/overview.md)
- [Sandcastles](/docs/sandcastles/overview.md)
- [MCP Servers](/docs/mcp-servers/overview.md)
- [Scheduled Triggers](/docs/scheduled-triggers/overview.md)
- [Agent Filesystem](/docs/agent-filesystem/overview.md)
- [Tool Policies](/docs/tool-policies/overview.md)
- [Workspace Permissions](/docs/workspace-permissions/overview.md)
- [Chat Sharing](/docs/chat-sharing/overview.md)

[Back to docs index](/docs.md)
