All Models
Qwen 3.8 Flash Next
Open-weight experimental preview of the Qwen4 architecture: hybrid-attention MoE (125B total, 6B active) with vision encoder for coding, agent tasks, and image and video understanding
Benchmarks
Available Providers (5)
| Provider | Model ID | Input Cost | Output Cost | Context | Max Output | Docs |
|---|---|---|---|---|---|---|
alibaba/qwen3.8-flash-next | $0.12/MTok | $0.40/MTok | 1.0M | 1.0M | ||
Qwen3.8-Flash-Next | $0.15/MTok | $0.47/MTok | 262.1K | 131.1K | ||
qwen3.8-flash-next | $0.20/MTok | $0.50/MTok | 262.1K | 262.1K | ||
qwen3.8-flash-next@eu | $0.20/MTok | $0.50/MTok | 262.1K | 262.1K | ||
qwen3.8-flash-next | $0.20/MTok | $0.50/MTok | 262.1K | 64K |
Capabilities
Reasoning
Tool Calling
Attachments
Open Weights
Structured Output