Skip to main content

Token usage

Monitor API usage across all keys and models.

Overview​

The Token Usage page displays four summary metrics, two trend charts, and a detailed usage table. Use the filters at the top to narrow results by time range or API key.

Summary cards​

Four cards display key metrics for the selected time range. Each card compares values against the previous period.

CardDescription
Total RequestsTotal number of API requests.
Total TokensTotal tokens consumed (input + output).
Active KeysNumber of API keys with activity in the selected period.
Avg LatencyAverage response latency in milliseconds.

Usage


Usage details​

The table lists per-row usage records with the following columns:

ColumnDescription
ModelThe model used.
API KeyThe API key used for the request.
RequestsNumber of requests.
Total TokensTotal tokens consumed.
PromptInput tokens in the prompt.
CompletionOutput tokens in the response.
Avg LatencyAverage response latency.
SuccessSuccess rate percentage.
ErrorsNumber of failed requests.
TimeTime window of the record.

The table supports pagination. Choose 10, 20, or 50 rows per page.

Next steps​

Manage API Keys

Create and manage credentials for deployed model endpoints.

Check Alerts

Filter, acknowledge, and resolve system alerts.