Vultr model API pricing

Vultr nvidia/Nemotron-3-Nano-Omni-30B-A3B-Reasoning-BF16 API pricing

Vultr nvidia/Nemotron-3-Nano-Omni-30B-A3B-Reasoning-BF16 API pricing is $0.130 per 1M input tokens and $0.380 per 1M output tokens in the ByteCosts provider pricing index. For an equal 1M input plus 1M output token plan, the listed token cost is $0.510. Context window: -. No source-backed prompt cache read/write price is published for this record. The confidence grade is A+: Read from the provider's official pricing page. The source was last checked 2026-07-19. At a reference workload of 1M input tokens and 250K output tokens per month, the listed token cost is $0.225 / month. This is a planning estimate; real invoices can differ because taxes, discounts, commitments, gateway fees, and non-token charges are excluded.

Open the pricing index - Vultr / nvidia/Nemotron-3-Nano-Omni-30B-A3B-Reasoning-BF16 →

Pricing snapshot

Vultr nvidia/Nemotron-3-Nano-Omni-30B-A3B-Reasoning-BF16 pricing fields
FieldValue
ProviderVultr
Modelnvidia/Nemotron-3-Nano-Omni-30B-A3B-Reasoning-BF16
Input / 1M tokens$0.130
Output / 1M tokens$0.380
Cache read / 1M tokens-
Cache write / 1M tokens-
Cache write 1h / 1M tokens-
Context window-
Reference monthly workload$0.225 / month

The reference workload is 1M input tokens plus 250K output tokens per month. It excludes taxes, discounts, gateway fees, and non-token charges.

Reference workload examples

Each example is pure arithmetic from this committed record's input and output token prices. The assumptions are shown inline.

Vultr nvidia/Nemotron-3-Nano-Omni-30B-A3B-Reasoning-BF16 monthly workload examples
ExampleAssumed monthly tokensEstimated token cost
Light reference1M input + 250K output tokens/month$0.225 / month
Production reference10M input + 2.5M output tokens/month$2.25 / month
Heavy reference100M input + 25M output tokens/month$22.50 / month

Examples exclude taxes, discounts, commitments, gateway fees, cache effects, and non-token charges unless those charges are part of the committed token price record.

When this model is cost-effective

nvidia/Nemotron-3-Nano-Omni-30B-A3B-Reasoning-BF16 ranks 2 of 9 Vultr records by combined list price ($0.510 for 1M input plus 1M output tokens). That puts it in the lower-priced third of this provider's committed records.

No source-backed cache-read discount is committed for this record, so the page does not assume cache savings.

No context window is committed for this record, so there is no long-context cost signal to apply.

Cheaper alternatives

These records have a lower committed input+output list price than this page's record. Same-provider records and same-model cross-provider records are preferred when available.

Vultr nvidia/Nemotron-3-Nano-Omni-30B-A3B-Reasoning-BF16 lower-priced alternatives
RecordInput / 1MOutput / 1MCombined / 1M + 1MDetail page
Vultr nvidia/Llama-3.1-Nemotron-Safety-Guard-8B-v3$0.010$0.010$0.020/tools/ai-provider-pricing/vultr--nvidia-llama-3.1-nemotron-safety-guard-8b-v3
Z.AI GLM-4.7-Flash$0$0$0/tools/ai-provider-pricing/zai--glm-4.7-flash
Z.AI GLM-4.5-Flash$0$0$0/tools/ai-provider-pricing/zai--glm-4.5-flash

Evidence and freshness

ByteCosts grades this row A+. Read from the provider's official pricing page. The record was first tracked on 2026-05-30 and last checked on 2026-07-19.

The source URL is kept with the record so readers can verify the provider page before making a purchasing or architecture decision. Prices change often, so quote the last-checked date with any copied number.

Watch-outs

  • no prompt caching

Price history

Vultr nvidia/Nemotron-3-Nano-Omni-30B-A3B-Reasoning-BF16 tracked changes
WhenFieldWasNowChange
2026-05-30trackingnot trackedcurrent rowno changes recorded

Frequently asked questions

Is nvidia/Nemotron-3-Nano-Omni-30B-A3B-Reasoning-BF16 free to use?

This ByteCosts row is not marked free. Use the listed token rates and verify the provider source for discounts, commitments, and account-specific terms.

Does this page include all possible charges?

No. The page focuses on source-backed token pricing and known cache/context fields. It does not include taxes, enterprise commitments, gateway markups, storage, retrieval, fine-tuning, or platform fees unless those charges are present in the source-backed row.

Vultr nvidia/Nemotron-3-Nano-Omni-30B-A3B-Reasoning-BF16 API pricing. ByteCosts. Updated 2026-07-19. https://bytecosts.com/tools/ai-provider-pricing/vultr--nvidia-nemotron-3-nano-omni-30b-a3b-reasoning-bf16/

Sources

Machine-readable