TheToolRadar

Llama 3.1 Nemotron 70B Instruct API Cost Calculator

Estimate your monthly API costs for Llama 3.1 Nemotron 70B Instruct from Nvidia.

Pricing per 1M tokens

Input
$1.20
Output
$1.20
131.1K context16.4K max outputFunction callingStreaming

Pricing last verified: Mar 21, 2026

30-Day Price History

Solid line: input price  ·  Dashed line: output price

Estimate Your Costs

What are you building?

Cost Breakdown

Input cost/request$0.000600
Output cost/request$0.000360
Cost per request$0.000960
Daily cost (1,000 requests)$0.9600
Monthly estimate$28.80
Annual estimate$345.60

Compare with Similar Models

ModelInput/1MOutput/1M
Nemotron Nano 12B 2 VL
Nvidia
$0.2000$0.6000
Nemotron 3 Super
Nvidia
$0.1000$0.5000
Llama 3.3 Nemotron Super 49B V1.5
Nvidia
$0.1000$0.4000
Nemotron 3 Nano 30B A3B
Nvidia
$0.0500$0.2000
Nemotron Nano 9B V2
Nvidia
$0.0400$0.1600

Frequently Asked Questions

How much does Llama 3.1 Nemotron 70B Instruct cost per token?

Llama 3.1 Nemotron 70B Instruct is priced at $1.20 per 1M input tokens and $1.20 per 1M output tokens. Use the calculator above to estimate your monthly costs based on your usage volume.

What is the context window for Llama 3.1 Nemotron 70B Instruct?

Llama 3.1 Nemotron 70B Instruct has a context window of 131.1K tokens (131,072 tokens). The maximum output length is 16.4K tokens.

Does Llama 3.1 Nemotron 70B Instruct support function calling (tool use)?

Yes, Llama 3.1 Nemotron 70B Instruct supports function calling (also called tool use). You can define custom functions that the model can invoke as part of its responses.

How does Llama 3.1 Nemotron 70B Instruct compare to similar models?

Llama 3.1 Nemotron 70B Instruct costs $1.20 input and $1.20 output per 1M tokens. Nemotron Nano 9B V2 is less expensive at $0.0400 input and $0.1600 output per 1M tokens. See the "Compare with Similar Models" table above, or visit the Nemotron Nano 9B V2 page at /tools/cost-calculator/nvidia/nemotron-nano-9b-v2.