Model usage and scheduled jobs, with 7-day and 30-day views.
Generated 2026-09-02
Figures and costs on this page are best-effort estimates from the machine that built it.
Per-day session token totals (agent JSONL) and cron job run token totals (runs store), aligned on UTC calendar days. Chart is inline SVG (no network); Y-axis is linear from 0 to the largest single-day total (model or cron). The table has exact counts.
Generated 2026-09-02 ยท
Source: [redacted path] ยท
Non-xAI token columns = session logs in the window (parseable timestamp). xAI: tokens + $/M from invoice preview (billing cycle); cost from POST โฆ/usage. Configure credentials (local): [REDACTED_AUTH] / XAI_TEAM_ID.
| Model | Provider | $/M In | $/M Out | Requests | Prompt Tokens | Compl. Tokens | Total Tokens | Est. Cost |
|---|---|---|---|---|---|---|---|---|
| deepseek/deepseek-v4-flash deepseek-v4-flash. Cache hit: $0.07/M. Max 8K output. | DeepSeek | $0.27 | $1.10 | 212 | 9,156,801 | 220,994 | 9,377,795 | $0.6950 |
| deepseek/deepseek-v4-pro deepseek-v4-pro. Thinking tokens count toward output cost. | DeepSeek | $0.55 | $2.19 | 31 | 1,538,347 | 68,543 | 1,606,890 | $0.1452 |
| google/gemini-2.5-flash Standard context (<200k). Thinking-mode output is $3.50/M. | Google AI Studio | โ | โ | 75 | 1,101,055 | 5,907 | 776,870 | $0.0000 |
| xai/grok-4-1-fast-reasoning Grok 4.1 Fast (reasoning + non-reasoning). 2M token context. | xAI | $0.20 | $0.50 | 0 | 0 | 0 | 0 | $0.0000 |
| nvidia/minimax-m2.7 MiniMax official pay-as-you-go pricing for M2.7 standard. | MiniMax via NVIDIA NIM | โ | โ | 5 | 0 | 0 | 0 | $0.0000 |
| nvidia/deepseek-v4-pro DeepSeek V4-Pro: 1.6T params, 49B active. NVIDIA NIM endpoint. | NVIDIA NIM (DeepSeek V4 Pro) | โ | โ | 0 | 0 | 0 | 0 | $0.0000 |
| nvidia/deepseek-v4-flash DeepSeek published pricing for V4-Flash. NVIDIA NIM endpoint. | NVIDIA NIM (DeepSeek V4 Flash) | โ | โ | 5 | 0 | 0 | 0 | $0.0000 |
| groq-main/llama-3.3-70b-versatile Groq LPU inference. 128K context, up to 33K output. | Groq | โ | โ | 16 | 0 | 0 | 0 | $0.0000 |
| cerebras-8b/llama-3.1-8b Cerebras Llama 3.1 8B. Very fast inference on Cerebras hardware. | Cerebras | โ | โ | 0 | 0 | 0 | 0 | $0.0000 |
| TOTAL | 344 | 11,796,203 | 295,444 | 11,761,555 | $0.8402 | |||
| Model | Provider | $/M In | $/M Out | Requests | Prompt Tokens | Compl. Tokens | Total Tokens | Est. Cost |
|---|---|---|---|---|---|---|---|---|
| deepseek/deepseek-v4-flash deepseek-v4-flash. Cache hit: $0.07/M. Max 8K output. | DeepSeek | $0.27 | $1.10 | 514 | 18,850,332 | 456,932 | 19,307,264 | $1.1790 |
| deepseek/deepseek-v4-pro deepseek-v4-pro. Thinking tokens count toward output cost. | DeepSeek | $0.55 | $2.19 | 88 | 3,887,838 | 155,514 | 4,043,352 | $0.3288 |
| google/gemini-2.5-flash Standard context (<200k). Thinking-mode output is $3.50/M. | Google AI Studio | โ | โ | 87 | 1,136,201 | 10,939 | 817,048 | $0.0000 |
| xai/grok-4-1-fast-reasoning Grok 4.1 Fast (reasoning + non-reasoning). 2M token context. | xAI | $0.20 | $0.50 | 0 | 0 | 0 | 0 | $0.0000 |
| nvidia/minimax-m2.7 MiniMax official pay-as-you-go pricing for M2.7 standard. | MiniMax via NVIDIA NIM | โ | โ | 6 | 0 | 0 | 0 | $0.0000 |
| nvidia/deepseek-v4-pro DeepSeek V4-Pro: 1.6T params, 49B active. NVIDIA NIM endpoint. | NVIDIA NIM (DeepSeek V4 Pro) | โ | โ | 0 | 0 | 0 | 0 | $0.0000 |
| nvidia/deepseek-v4-flash DeepSeek published pricing for V4-Flash. NVIDIA NIM endpoint. | NVIDIA NIM (DeepSeek V4 Flash) | โ | โ | 6 | 0 | 0 | 0 | $0.0000 |
| groq-main/llama-3.3-70b-versatile Groq LPU inference. 128K context, up to 33K output. | Groq | โ | โ | 16 | 0 | 0 | 0 | $0.0000 |
| cerebras-8b/llama-3.1-8b Cerebras Llama 3.1 8B. Very fast inference on Cerebras hardware. | Cerebras | โ | โ | 0 | 0 | 0 | 0 | $0.0000 |
| TOTAL | 717 | 23,874,371 | 623,385 | 24,167,664 | $1.5078 | |||
|