Provider pricing
xAI pricing: API cost per model, user, and month
xAI lists 7 priced models in the ByteCosts index, with Grok Build 0.1 at about $1.30 per million tokens on ByteCosts’ 70% input / 30% output blend. For most AI apps the bill is driven by output tokens, retry rate, and prompt-cache hit rate far more than the headline input price, so xAI is cost-effective when your workload is input-heavy (RAG, classification) or can cache a large shared prefix. Compare xAI against alternatives on a real workload, seats, requests, and token mix, before committing to a model or a subscription, because a cheaper per-token price can still lose once power users and long outputs are priced in.
Estimate your monthly cost - Turn your usage into xAI spend →
xAI pricing at a glance
- Cheapest tracked chat model: Grok Build 0.1 at $1.30 per 1M tokens (blended)
- Flagship: Grok 4.5 at $3.20 per 1M tokens (blended)
- Last recorded price event: Grok Build 0.1 priority input on Aug 22, 2026
- 7 priced models in the index, updated August 22, 2026
xAI model prices
Per-million-token list prices for current xAI models, from the ByteCosts pricing index. Output tokens usually cost several times more than input.
| Model | Input | Output | Context |
|---|---|---|---|
| Grok Build 0.1 | $1.00 | $2.00 | 256K |
| Grok 4.20 (Non-Reasoning) | $1.25 | $2.50 | 1M |
| Grok 4.20 (Reasoning) | $1.25 | $2.50 | 1M |
| Grok 4.20 Multi-Agent | $1.25 | $2.50 | 1M |
| Grok 4.3 | $1.25 | $2.50 | 1M |
| Grok 4.5 | $2.00 | $6.00 | 500K |
| Grok 4.6 | $2.00 | $6.00 | 500K |
xAI subscription plans
Flat-rate xAI plans billed per month, separate from the per-token API pricing above. List prices; annual or quarterly billing can lower the effective rate. Verify the current price with xAI before subscribing.
| Plan | Monthly | Annual (effective) | Includes |
|---|---|---|---|
| Grok SuperGrok | $30/mo | - | Grok Build / coding tools. Not a Grok Bot grant. |
| Grok Business | $30/mo | - | Team Grok workspace. SuperGrok Team/Enterprise do not grant Grok Bot on Cursor. |
| Grok SuperGrok Plus | $100/mo | - | Grok Bot plus Grok Build. A SuperGrok Plus link meters Grok Bot on the Cursor account and does not change the Cursor plan. |
Recent xAI price events
Source-backed price events recorded for xAI in the ByteCosts ledger, newest first. Every event below is a first-time listing, the ledger started tracking that rate; the model itself may be older.
| Model | Rate | Price | Recorded |
|---|---|---|---|
| Grok Build 0.1 | priority input | $2.00 | Aug 22, 2026 |
| Grok Build 0.1 | priority output | $4.00 | Aug 22, 2026 |
| Grok 4.3 | batch input | $1.00 | Aug 4, 2026 |
| Grok 4.3 | batch output | $2.00 | Aug 4, 2026 |
| Grok 4.3 | priority input | $2.50 | Aug 4, 2026 |
| Grok 4.3 | priority output | $5.00 | Aug 4, 2026 |
| Grok 4.3 | cached input | $0.200 | Jun 1, 2026 |
| Grok 4.3 | cache read | $0.200 | Jun 1, 2026 |
| Grok Build 0.1 | cached input | $0.200 | Jun 1, 2026 |
| Grok Build 0.1 | cache read | $0.200 | Jun 1, 2026 |
| Imagine API - Image | per-image | $0.020 / image | Jun 1, 2026 |
When xAI is cost-effective
Token price alone does not decide cost. The variables that move an AI bill are: how many output tokens each call emits (output is the expensive side), how often calls retry, what fraction of the input prefix can be served from prompt cache, and how heavily your top 1% of users use the product.
xAI tends to win when your workload is input-heavy with short outputs, when you can cache a large shared system prompt or document context, or when a smaller xAI model clears your quality bar. It tends to lose when outputs are long and uncached, where a cheaper-per-output model compounds the saving across millions of calls.
Limitations before production billing decisions
Treat ByteCosts calculations as planning estimates, not final billing totals. Real invoices can differ because token mix, retry rate, cache hit rate, rate limits, taxes, gateway fees, regional pricing, and negotiated discounts change the effective cost.
Verify the provider source before production billing decisions, then compare the estimate with your own logs or invoice once production traffic is live.
Lower-blend alternatives to xAI
On ByteCosts’ 70% input / 30% output chat blend, these providers screen lower than xAI. Treat this as a shortlist, not a decision, check the selected model against your real token mix:
- Mistral: $0.040 blended via Ministral 3B (latest), 28 priced models.
- Cohere: $0.071 blended via Command R7B, 8 priced models.
- Alibaba Qwen: $0.095 blended via Qwen Turbo, 46 priced models.
Frequently asked questions
How much does xAI cost per million tokens?
xAI‘s low-cost tracked chat row is Grok Build 0.1 at about $1.30 per million tokens on ByteCosts’ 70% input / 30% output blend. That same model’s component rates are $1.00 input and $2.00 output; flagship models cost more. See the price table above for current per-model rates.
Is xAI cheaper than the alternatives?
It depends on the workload. xAI can be cheaper for input-heavy or cacheable workloads even when its headline price is higher, because output volume and cache hit rate move the bill more than the per-token rate. Use the ByteCosts calculators to compare on your real traffic.
Where does ByteCosts get xAI prices?
Prices come from xAI‘s official pricing/docs pages, normalized into the ByteCosts pricing index and dated. Each record carries a confidence grade and a source link. Prices are list prices and exclude negotiated or volume discounts. Verify the provider source before production billing decisions.
xAI pricing. ByteCosts. Updated August 22, 2026. https://bytecosts.com/pricing/xai/