Qwen3 Coder Flash
Best open-weights code model in its class. A strong Codestral / GPT-4 alternative for self-hosted IDE backends.
Qwen3 Coder Flash is a efficient AI model from Qwen. It costs $0.195 per million input tokens and $0.975 per million output tokens (blended $0.429/M), with a 1,000,000-token context window.
Profile inherited from upstream Qwen Coder (Qwen3 base) ↗ — this is a hosted variant of the same open-weights model.
- Best open-weights code at 32B
- 128K context
- Permissive license
- Self-hosted IDE backends
- Code review bots
Benchmarks
More from Qwen
See all 22 →Frequently asked questions
How much does Qwen3 Coder Flash cost?
Qwen3 Coder Flash costs $0.195 per million input tokens and $0.975 per million output tokens, for a blended reference rate of $0.429 per million tokens.
What is Qwen3 Coder Flash's context window?
Qwen3 Coder Flash supports up to 1,000,000 tokens of context in a single request.
What is Qwen3 Coder Flash best for?
Qwen3 Coder Flash is well suited to Best open-weights code at 32B, 128K context and Permissive license.
Who makes Qwen3 Coder Flash?
Qwen3 Coder Flash is developed and served by Qwen.