Llama 3.1 Nemotron 70B Instruct API Cost Calculator
Estimate your monthly API costs for Llama 3.1 Nemotron 70B Instruct from Nvidia.
Pricing per 1M tokens
Pricing last verified: Mar 21, 2026
30-Day Price History
Solid line: input price · Dashed line: output price
Estimate Your Costs
What are you building?
Cost Breakdown
Compare with Similar Models
| Model | Input/1M | Output/1M |
|---|---|---|
| Nemotron Nano 12B 2 VL Nvidia | $0.2000 | $0.6000 |
| Nemotron 3 Super Nvidia | $0.1000 | $0.5000 |
| Llama 3.3 Nemotron Super 49B V1.5 Nvidia | $0.1000 | $0.4000 |
| Nemotron 3 Nano 30B A3B Nvidia | $0.0500 | $0.2000 |
| Nemotron Nano 9B V2 Nvidia | $0.0400 | $0.1600 |
Frequently Asked Questions
How much does Llama 3.1 Nemotron 70B Instruct cost per token?
Llama 3.1 Nemotron 70B Instruct is priced at $1.20 per 1M input tokens and $1.20 per 1M output tokens. Use the calculator above to estimate your monthly costs based on your usage volume.
What is the context window for Llama 3.1 Nemotron 70B Instruct?
Llama 3.1 Nemotron 70B Instruct has a context window of 131.1K tokens (131,072 tokens). The maximum output length is 16.4K tokens.
Does Llama 3.1 Nemotron 70B Instruct support function calling (tool use)?
Yes, Llama 3.1 Nemotron 70B Instruct supports function calling (also called tool use). You can define custom functions that the model can invoke as part of its responses.
How does Llama 3.1 Nemotron 70B Instruct compare to similar models?
Llama 3.1 Nemotron 70B Instruct costs $1.20 input and $1.20 output per 1M tokens. Nemotron Nano 9B V2 is less expensive at $0.0400 input and $0.1600 output per 1M tokens. See the "Compare with Similar Models" table above, or visit the Nemotron Nano 9B V2 page at /tools/cost-calculator/nvidia/nemotron-nano-9b-v2.