Fireworks AI model API pricing
Fireworks AI Base Embeddings variant 2 API pricing
Fireworks AI Base Embeddings API pricing is $0.016 per 1M input tokens and - per 1M output tokens in the ByteCosts provider pricing index. A combined 1M input plus 1M output token estimate is not calculated because one of those prices is missing. Context window: 150M. Prompt cache prices: read $0.0080, write -, 1h write - per 1M tokens. The confidence grade is A+: Read from the provider's official pricing page. The source was last checked 2026-07-19. A reference monthly cost is not calculated because either input or output price is missing. This is a planning estimate; real invoices can differ because taxes, discounts, commitments, gateway fees, and non-token charges are excluded.
Open the pricing index - Fireworks AI / Base Embeddings →
Pricing snapshot
| Field | Value |
|---|---|
| Provider | Fireworks AI |
| Model | Base Embeddings |
| Input / 1M tokens | $0.016 |
| Output / 1M tokens | - |
| Cache read / 1M tokens | $0.0080 |
| Cache write / 1M tokens | - |
| Cache write 1h / 1M tokens | - |
| Context window | 150M |
| Reference monthly workload | not calculated |
The reference workload is 1M input tokens plus 250K output tokens per month. It excludes taxes, discounts, gateway fees, and non-token charges.
Reference workload examples
Each example is pure arithmetic from this committed record's input and output token prices. The assumptions are shown inline.
| Example | Assumed monthly tokens | Estimated token cost |
|---|---|---|
| Light reference | 1M input + 250K output tokens/month | not calculated |
| Production reference | 10M input + 2.5M output tokens/month | not calculated |
| Heavy reference | 100M input + 25M output tokens/month | not calculated |
Examples exclude taxes, discounts, commitments, gateway fees, cache effects, and non-token charges unless those charges are part of the committed token price record.
When this model is cost-effective
ByteCosts cannot rank Base Embeddings within Fireworks AI by combined input+output price because one of the token prices is missing.
Cached input reads are listed at $0.0080 per 1M tokens, 50% of the fresh input rate, so repeated cached prompts have a lower committed input-token cost than fresh prompts.
The committed context window is 150M; use that as the only long-context signal on this page, not as a quality or latency claim.
Cheaper alternatives
These records have a lower committed input+output list price than this page's record. Same-provider records and same-model cross-provider records are preferred when available.
| Record | Input / 1M | Output / 1M | Combined / 1M + 1M | Detail page |
|---|---|---|---|---|
| No lower-priced committed record | - | - | - | - |
Evidence and freshness
ByteCosts grades this row A+. Read from the provider's official pricing page. The record was first tracked on 2026-05-30 and last checked on 2026-07-19.
The source URL is kept with the record so readers can verify the provider page before making a purchasing or architecture decision. Prices change often, so quote the last-checked date with any copied number.
Watch-outs
- No hidden-cost flags are attached to this source-backed record.
Price history
| When | Field | Was | Now | Change |
|---|---|---|---|---|
| 2026-06-01 | embedding | new | $0.016 | new |
Frequently asked questions
Is Base Embeddings free to use?
This ByteCosts row is not marked free. Use the listed token rates and verify the provider source for discounts, commitments, and account-specific terms.
Does this page include all possible charges?
No. The page focuses on source-backed token pricing and known cache/context fields. It does not include taxes, enterprise commitments, gateway markups, storage, retrieval, fine-tuning, or platform fees unless those charges are present in the source-backed row.
Fireworks AI Base Embeddings variant 2 API pricing. ByteCosts. Updated 2026-07-19. https://bytecosts.com/tools/ai-provider-pricing/fireworks-ai--base-embeddings--dup2/