Qwen: Qwen3 8B
qwen/qwen3-8b
Description
Qwen3-8B is a dense 8.2B parameter causal language model from the Qwen3 series, designed for both reasoning-heavy tasks and efficient dialogue. It supports seamless switching between "thinking" mode for math, coding, and logical inference, and "non-thinking" mode for general conversation. The model is fine-tuned for instruction-following, agent integration, creative writing, and multilingual use across 100+ languages and dialects. It natively supports a 32K token context window and can extend to 131K tokens with YaRN scaling.
How this model compares
Overall covers the full catalog. By plan covers only models available on that tier (same rules as available models in your list). Position on min–average–max. Prices use the higher of prompt or completion per token, shown per 1M tokens.
Price (per 1M tokens)
Min
Max
This model
339 models in this groupPrice (per 1M tokens)
- Min
- $0.04
- Avg
- $12.395447
- Max
- $750.00
This model: $0.40 / 1M tokens
Context length (tokens)
Min
Max
This model
339 models in this groupContext length (tokens)
- Min
- 4,095 tokens
- Avg
- 379,884.782 tokens
- Max
- 10,000,000 tokens
This model: 131,072 tokens
Capabilities
Text → TextContext: 40,960 tokens
Input:
Text
Output:
Text