Alibaba

Qwen3.8 Flash NextOpen

Alibaba Qwen3.8 Flash Next experimental high-throughput multimodal preview model.

#163released Aug 18, 2026·Open Weights·Not disclosed
textimagevideo260K ctx
API providersLeaderboard
Intelligence
66.8 #163
AAQI 61.5
Output speed
168 t/s
TTFT 210 ms · ITL 6.0 ms/tok
Blended price (3:1)
$0.19 /1M
In $0.12/M · Out $0.4/M
Context
260K
≈ 650 A4 pages · max out 8192

Quality Domain Radar

6-Axis Frontier Evaluation vs Category Average & Flagship

Qwen3.8 Flash Next
Claude Mythos Preview (#1)
Average (68)
Coding (60.5)Mathematics (68)Hard Science (GPQA) (66)Agentic & Tool Use (63.5)Knowledge (MMLU) (64)Vision / Multimodal (66)

Radar axes are approximated from category scores where direct benchmark measurements are missing — see Methodology.

API deployment

Compare hosts & latency

Throughput, TTFT and cache pricing across Groq, Cerebras, Together, DeepInfra, Fireworks, Bedrock and Azure.

GroqCerebrasTogether AIDeepInfraFireworks AIAWS BedrockAzure AI
View providers matrix →

Scores by Source

Model Details

Release date
Aug 18, 2026
Parameters
Not disclosed
License
Open Weights
Context Window
260,000 tokens
Input Price
$0.12/M
Output Price
$0.4/M
Output Speed
168 t/s
Time to First Token
210 ms

Estimated monthly bill

Estimated monthly spend
$0.00
Calculating savings vs GPT-4o...

Cite & embed

GitHub Markdown badge
[![LLMPodium Rank](https://llmpodium.com/badge/qwen3-8-flash-next/rank.svg)](https://llmpodium.com/models/qwen3-8-flash-next)
BibTeX
@misc{llmpodium2026qwen38flashnext,
  title = {Benchmark Analysis and Intelligence Evaluation of Qwen3.8 Flash Next},
  author = {LLMPodium Research},
  year = {2026},
  url = {https://llmpodium.com/models/qwen3-8-flash-next}
}

More from Alibaba

Alternative Models

Closest alternatives to Qwen3.8 Flash Next based on capabilities, pricing, and benchmark scores.

0 / 4