Groq model API pricing
Groq Llama 4 Scout (17Bx16E) 128k API pricing
Groq Llama 4 Scout (17Bx16E) 128k API pricing is $0.110 per 1M input tokens and $0.340 per 1M output tokens in the ByteCosts provider pricing index. For an equal 1M input plus 1M output token plan, the listed token cost is $0.450. Context window: 128K. No source-backed prompt cache read/write price is published for this record. The confidence grade is A+: Read from the provider’s official pricing page. The source was last checked 2026-08-22. At a reference workload of 1M input tokens and 250K output tokens per month, the listed token cost is $0.195 / month. This is a planning estimate; real invoices can differ because taxes, discounts, commitments, gateway fees, and non-token charges are excluded.
Open the pricing index - Groq / Llama 4 Scout (17Bx16E) 128k →
Pricing snapshot
| Field | Value |
|---|---|
| Provider | Groq |
| Model | Llama 4 Scout (17Bx16E) 128k |
| Input / 1M tokens | $0.110 |
| Output / 1M tokens | $0.340 |
| Cache read / 1M tokens | - |
| Cache write / 1M tokens | - |
| Cache write 1h / 1M tokens | - |
| Context window | 128K |
| Reference monthly workload | $0.195 / month |
The reference workload is 1M input tokens plus 250K output tokens per month. It excludes taxes, discounts, gateway fees, and non-token charges.
Reference workload examples
Each example is pure arithmetic from this committed record’s input and output token prices. The assumptions are shown inline.
| Example | Assumed monthly tokens | Estimated token cost |
|---|---|---|
| Light reference | 1M input + 250K output tokens/month | $0.195 / month |
| Production reference | 10M input + 2.5M output tokens/month | $1.95 / month |
| Heavy reference | 100M input + 25M output tokens/month | $19.50 / month |
Examples exclude taxes, discounts, commitments, gateway fees, cache effects, and non-token charges unless those charges are part of the committed token price record.
When this model is cost-effective
Llama 4 Scout (17Bx16E) 128k ranks 5 of 10 Groq records by combined list price ($0.450 for 1M input plus 1M output tokens). That puts it in the middle third of this provider's committed records.
No source-backed cache-read discount is committed for this record, so the page does not assume cache savings.
The committed context window is 128K; use that as the only long-context signal on this page, not as a quality or latency claim.
Cheaper alternatives
These records have a lower committed input+output list price than this page’s record. Same-provider records and same-model cross-provider records are preferred when available.
| Record | Input / 1M | Output / 1M | Combined / 1M + 1M | Detail page |
|---|---|---|---|---|
| Groq Llama 3.1 8B Instant 128k | $0.050 | $0.080 | $0.130 | /tools/ai-provider-pricing/groq--llama-3.1-8b-instant-128k |
| Groq GPT OSS 20B 128k | $0.075 | $0.300 | $0.375 | /tools/ai-provider-pricing/groq--gpt-oss-20b-128k |
| Groq GPT OSS Safeguard 20B | $0.075 | $0.300 | $0.375 | /tools/ai-provider-pricing/groq--gpt-oss-safeguard-20b |
Evidence and freshness
ByteCosts grades this row A+. Read from the provider’s official pricing page. The record was first tracked on 2026-05-30 and last checked on 2026-08-22.
The source URL is kept with the record so readers can verify the provider page before making a purchasing or architecture decision. Prices change often, so quote the last-checked date with any copied number.
Watch-outs
- no prompt caching
Price history
| When | Field | Was | Now | Change |
|---|---|---|---|---|
| 2026-08-22 | batchInput | new | $0.055 | new |
| 2026-08-22 | batchOutput | new | $0.170 | new |
Frequently asked questions
Is Llama 4 Scout (17Bx16E) 128k free to use?
This ByteCosts row is not marked free. Use the listed token rates and verify the provider source for discounts, commitments, and account-specific terms.
Does this page include all possible charges?
No. The page focuses on source-backed token pricing and known cache/context fields. It does not include taxes, enterprise commitments, gateway markups, storage, retrieval, fine-tuning, or platform fees unless those charges are present in the source-backed row.
Groq Llama 4 Scout (17Bx16E) 128k API pricing. ByteCosts. Updated 2026-08-22. https://bytecosts.com/tools/ai-provider-pricing/groq--llama-4-scout-17bx16e-128k/