Fireworks AI model API pricing

Fireworks AI Base Embeddings API pricing

Fireworks AI Base Embeddings API pricing is $0.0080 per 1M input tokens and - per 1M output tokens in the ByteCosts provider pricing index. A combined 1M input plus 1M output token estimate is not calculated because one of those prices is missing. Context window: 150M. Prompt cache prices: read $0.0040, write -, 1h write - per 1M tokens. The confidence grade is A+: Read from the provider's official pricing page. The source was last checked 2026-07-19. A reference monthly cost is not calculated because either input or output price is missing. This is a planning estimate; real invoices can differ because taxes, discounts, commitments, gateway fees, and non-token charges are excluded.

Open the pricing index - Fireworks AI / Base Embeddings →

Pricing snapshot

Fireworks AI Base Embeddings pricing fields
FieldValue
ProviderFireworks AI
ModelBase Embeddings
Input / 1M tokens$0.0080
Output / 1M tokens-
Cache read / 1M tokens$0.0040
Cache write / 1M tokens-
Cache write 1h / 1M tokens-
Context window150M
Reference monthly workloadnot calculated

The reference workload is 1M input tokens plus 250K output tokens per month. It excludes taxes, discounts, gateway fees, and non-token charges.

Reference workload examples

Each example is pure arithmetic from this committed record's input and output token prices. The assumptions are shown inline.

Fireworks AI Base Embeddings monthly workload examples
ExampleAssumed monthly tokensEstimated token cost
Light reference1M input + 250K output tokens/monthnot calculated
Production reference10M input + 2.5M output tokens/monthnot calculated
Heavy reference100M input + 25M output tokens/monthnot calculated

Examples exclude taxes, discounts, commitments, gateway fees, cache effects, and non-token charges unless those charges are part of the committed token price record.

When this model is cost-effective

ByteCosts cannot rank Base Embeddings within Fireworks AI by combined input+output price because one of the token prices is missing.

Cached input reads are listed at $0.0040 per 1M tokens, 50% of the fresh input rate, so repeated cached prompts have a lower committed input-token cost than fresh prompts.

The committed context window is 150M; use that as the only long-context signal on this page, not as a quality or latency claim.

Cheaper alternatives

These records have a lower committed input+output list price than this page's record. Same-provider records and same-model cross-provider records are preferred when available.

Fireworks AI Base Embeddings lower-priced alternatives
RecordInput / 1MOutput / 1MCombined / 1M + 1MDetail page
No lower-priced committed record----

Evidence and freshness

ByteCosts grades this row A+. Read from the provider's official pricing page. The record was first tracked on 2026-05-30 and last checked on 2026-07-19.

The source URL is kept with the record so readers can verify the provider page before making a purchasing or architecture decision. Prices change often, so quote the last-checked date with any copied number.

Watch-outs

  • No hidden-cost flags are attached to this source-backed record.

Price history

Fireworks AI Base Embeddings tracked changes
WhenFieldWasNowChange
2026-06-01embeddingnew$0.0080new

Frequently asked questions

Is Base Embeddings free to use?

This ByteCosts row is not marked free. Use the listed token rates and verify the provider source for discounts, commitments, and account-specific terms.

Does this page include all possible charges?

No. The page focuses on source-backed token pricing and known cache/context fields. It does not include taxes, enterprise commitments, gateway markups, storage, retrieval, fine-tuning, or platform fees unless those charges are present in the source-backed row.

Fireworks AI Base Embeddings API pricing. ByteCosts. Updated 2026-07-19. https://bytecosts.com/tools/ai-provider-pricing/fireworks-ai--base-embeddings/

Sources

Machine-readable