Operational

Usage

Token consumption against your current plan allowance, broken down by model.

Requests128,310last 90 days
Total tokens85,282,38652,875,079 in / 32,407,307 out
Spend$361.26billed usage
Mean latency621 ms
Success rate98.28%0.4%

Plan allowance

Monthly plan · renews Oct 18, 2026

41%of plan allowance used
412,884,000 of 1,000,000,000 tokens · 28 days remaining in this period

Usage

Tokens over the last 7 days

Usage by model

Input and output token consumption across the last 90 days.

ModelRequestsInput tokensOutput tokensMean latencySuccessSpend
Claude Sonnet 5claude-sonnet-53,7951,153,680507,619869 ms99.5%$7.97
DeepSeek V4.1 Flashdeepseek-v4.1-flash1,7771,254,562552,007571 ms97.6%$3.73
GLM 5.2glm-5.21,7281,016,064447,0681,159 ms99.8%$3.63
GLM 5.3 Flashglm-5.3-flash1,037599,386263,730946 ms97.6%$2.18
GPT 5.6 Lunagpt-5.6-luna720220,32096,941696 ms97.3%$1.51
DeepSeek V4 Prodeepseek-v4-pro511348,502153,341348 ms97.6%$1.07
  • Claude Sonnet 5claude-sonnet-5 · 3,795 requests · 1,661,299 tokens · 99.5% success
  • DeepSeek V4.1 Flashdeepseek-v4.1-flash · 1,777 requests · 1,806,569 tokens · 97.6% success
  • GLM 5.2glm-5.2 · 1,728 requests · 1,463,132 tokens · 99.8% success
  • GLM 5.3 Flashglm-5.3-flash · 1,037 requests · 863,116 tokens · 97.6% success
  • GPT 5.6 Lunagpt-5.6-luna · 720 requests · 317,261 tokens · 97.3% success
  • DeepSeek V4 Prodeepseek-v4-pro · 511 requests · 501,843 tokens · 97.6% success

How usage is measured

What counts toward the allowance and what does not.

  • Prompt tokens include the system message, history and tool definitions sent with the request.
  • Completion tokens include reasoning tokens when a model emits them.
  • Requests rejected before reaching a provider — bad key, unknown model — do not consume the allowance.
  • When the allowance is exhausted, requests return 429 until the period renews.