Qwen3 VL 8B Thinking API Cost Calculator
Estimate your monthly API costs for Qwen3 VL 8B Thinking from Qwen.
Pricing per 1M tokens
Pricing last verified: Mar 21, 2026
30-Day Price History
Solid line: input price · Dashed line: output price
Estimate Your Costs
What are you building?
Cost Breakdown
Compare with Similar Models
| Model | Input/1M | Output/1M |
|---|---|---|
| Qwen3.5-35B-A3B Qwen | $0.1625 | $1.30 |
| Qwen2.5 VL 72B Instruct Qwen | $0.8000 | $0.8000 |
| Qwen3 235B A22B Thinking 2507 Qwen | $0.1495 | $1.50 |
| Qwen2.5 Coder 32B Instruct Qwen | $0.6600 | $1.00 |
| Qwen3 VL 30B A3B Thinking Qwen | $0.1300 | $1.56 |
Frequently Asked Questions
How much does Qwen3 VL 8B Thinking cost per token?
Qwen3 VL 8B Thinking is priced at $0.1170 per 1M input tokens and $1.37 per 1M output tokens. Use the calculator above to estimate your monthly costs based on your usage volume.
What is the context window for Qwen3 VL 8B Thinking?
Qwen3 VL 8B Thinking has a context window of 131.1K tokens (131,072 tokens). The maximum output length is 32.8K tokens.
Does Qwen3 VL 8B Thinking support image inputs (vision)?
Yes, Qwen3 VL 8B Thinking supports vision/image inputs. You can send images as part of your prompts.
Does Qwen3 VL 8B Thinking support function calling (tool use)?
Yes, Qwen3 VL 8B Thinking supports function calling (also called tool use). You can define custom functions that the model can invoke as part of its responses.
How does Qwen3 VL 8B Thinking compare to similar models?
Qwen3 VL 8B Thinking costs $0.1170 input and $1.37 output per 1M tokens. Qwen3.5-35B-A3B is less expensive at $0.1625 input and $1.30 output per 1M tokens. See the "Compare with Similar Models" table above, or visit the Qwen3.5-35B-A3B page at /tools/cost-calculator/qwen/qwen3-5-35b-a3b.