TheToolRadar

Qwen3 VL 8B Thinking API Cost Calculator

Estimate your monthly API costs for Qwen3 VL 8B Thinking from Qwen.

Pricing per 1M tokens

Input
$0.1170
Output
$1.37
131.1K context32.8K max outputVisionFunction callingStreaming

Pricing last verified: Mar 21, 2026

30-Day Price History

Solid line: input price  ·  Dashed line: output price

Estimate Your Costs

What are you building?

Cost Breakdown

Input cost/request$0.000059
Output cost/request$0.000409
Cost per request$0.000468
Daily cost (1,000 requests)$0.4680
Monthly estimate$14.04
Annual estimate$168.48

Compare with Similar Models

ModelInput/1MOutput/1M
Qwen3.5-35B-A3B
Qwen
$0.1625$1.30
Qwen2.5 VL 72B Instruct
Qwen
$0.8000$0.8000
Qwen3 235B A22B Thinking 2507
Qwen
$0.1495$1.50
Qwen2.5 Coder 32B Instruct
Qwen
$0.6600$1.00
Qwen3 VL 30B A3B Thinking
Qwen
$0.1300$1.56

Frequently Asked Questions

How much does Qwen3 VL 8B Thinking cost per token?

Qwen3 VL 8B Thinking is priced at $0.1170 per 1M input tokens and $1.37 per 1M output tokens. Use the calculator above to estimate your monthly costs based on your usage volume.

What is the context window for Qwen3 VL 8B Thinking?

Qwen3 VL 8B Thinking has a context window of 131.1K tokens (131,072 tokens). The maximum output length is 32.8K tokens.

Does Qwen3 VL 8B Thinking support image inputs (vision)?

Yes, Qwen3 VL 8B Thinking supports vision/image inputs. You can send images as part of your prompts.

Does Qwen3 VL 8B Thinking support function calling (tool use)?

Yes, Qwen3 VL 8B Thinking supports function calling (also called tool use). You can define custom functions that the model can invoke as part of its responses.

How does Qwen3 VL 8B Thinking compare to similar models?

Qwen3 VL 8B Thinking costs $0.1170 input and $1.37 output per 1M tokens. Qwen3.5-35B-A3B is less expensive at $0.1625 input and $1.30 output per 1M tokens. See the "Compare with Similar Models" table above, or visit the Qwen3.5-35B-A3B page at /tools/cost-calculator/qwen/qwen3-5-35b-a3b.