TheToolRadar

Llama 3.3 Nemotron Super 49B V1.5 API Cost Calculator

Estimate your monthly API costs for Llama 3.3 Nemotron Super 49B V1.5 from Nvidia.

Pricing per 1M tokens

Input
$0.1000
Output
$0.4000
131.1K contextFunction callingStreaming

Pricing last verified: Mar 21, 2026

30-Day Price History

Solid line: input price  ·  Dashed line: output price

Estimate Your Costs

What are you building?

Cost Breakdown

Input cost/request$0.000050
Output cost/request$0.000120
Cost per request$0.000170
Daily cost (1,000 requests)$0.1700
Monthly estimate$5.10
Annual estimate$61.20

Compare with Similar Models

ModelInput/1MOutput/1M
Nemotron 3 Super
Nvidia
$0.1000$0.5000
Nemotron 3 Nano 30B A3B
Nvidia
$0.0500$0.2000
Nemotron Nano 9B V2
Nvidia
$0.0400$0.1600
Nemotron Nano 12B 2 VL
Nvidia
$0.2000$0.6000
Nemotron 3 Nano 30B A3B (free)
Nvidia
$0.00$0.00

Frequently Asked Questions

How much does Llama 3.3 Nemotron Super 49B V1.5 cost per token?

Llama 3.3 Nemotron Super 49B V1.5 is priced at $0.1000 per 1M input tokens and $0.4000 per 1M output tokens. Use the calculator above to estimate your monthly costs based on your usage volume.

What is the context window for Llama 3.3 Nemotron Super 49B V1.5?

Llama 3.3 Nemotron Super 49B V1.5 has a context window of 131.1K tokens (131,072 tokens).

Does Llama 3.3 Nemotron Super 49B V1.5 support function calling (tool use)?

Yes, Llama 3.3 Nemotron Super 49B V1.5 supports function calling (also called tool use). You can define custom functions that the model can invoke as part of its responses.

How does Llama 3.3 Nemotron Super 49B V1.5 compare to similar models?

Llama 3.3 Nemotron Super 49B V1.5 costs $0.1000 input and $0.4000 output per 1M tokens. Nemotron 3 Nano 30B A3B (free) is less expensive at $0.00 input and $0.00 output per 1M tokens. See the "Compare with Similar Models" table above, or visit the Nemotron 3 Nano 30B A3B (free) page at /tools/cost-calculator/nvidia/nemotron-3-nano-30b-a3b-free.