Alibaba Just Released Qwen3.8-Flash: “An Early Preview of the Architecture in Qwen4”
The open-weight model uses 6 billion active parameters per token and is meant to preview the Qwen 4 architecture for developers and enterprise teams.
9 Articles
9 Articles
Alibaba just released Qwen3.8-Flash: “An early preview of the architecture in Qwen4”
Alibaba this week unveiled Qwen3.8-Flash, an open-weight, multimodal Mixture-of-Experts (MoE) model. Hot on the heels of Qwen 3.8 Max, which arrived at the start of the month, this 125-billion-parameter AI model is offered as a prelude to Qwen 4. It is positioned as both a performance and value-for-money play. As such, it is claimed to have “superior capabilities in coding and office tasks” and an optimal balance among capability, latency, and …
Qwen3.8-Flash Matches DeepSeek V4 Pro at a Quarter of the Price
Alibaba’s Qwen team released Qwen3.8-Flash on August 26, 2026, a multimodal mixture-of-experts model that activates only 6 billion parameters per token despite carrying 125 billion total. On paper, it is an efficiency play. In practice, it beats DeepSeek V4 Pro on SWE-bench Pro (62.5 vs. 55.4) while costing roughly a quarter of the price, and it lands within striking distance of Claude Opus 5 on coding benchmarks at 33x lower input cost. The rel…
Qwen 3.8-Flash-Next From Alibaba: The First Look at Qwen 4
Alibaba’s Qwen team didn’t wait for a big stage. On August 26, 2026, they quietly released Qwen 3.8-Flash-Next — a 125 billion parameter AI model that’s designed to run on hardware most developers already own. It’s not the finished Qwen 4. It’s the architecture preview, the sneak peek that tells developers what’s coming and lets them start building now. Here’s what’s inside, what’s missing, and why it matters. What Qwen 3.8-Flash-Next Actually I…
Alibaba’s QwenWork adds Qwen3.8-Flash with a new Standard mode · TechNode
Alibaba’s QwenWork has added Qwen3.8-Flash and launched a new Standard mode that is available to all users. The service will now offer Standard and Advanced modes for different levels of office tasks. Qwen says about 95% of daily tasks can be handled by the Standard mode, while more complex tasks require the Advanced mode. In tests cited by the company, the new mode increased single-task generation speed by about 100% and reduced token consumpti…
GLM-5.3-Flash vs Qwen3.8-Flash-Next: Two Chinese AI Labs Independently Converge on the Same Model Architecture
Z.ai and Qwen independently shipped near-identical architectures: 3:1 linear hybrids, compressed indexers, gated residuals, and Muon training. The post GLM-5.3-Flash vs Qwen3.8-Flash-Next: Two Chinese AI Labs Independently Converge on the Same Model Architecture appeared first on MarkTechPost.
Alibaba Cloud has launched Qwen3.8-Flash-Next, a multimodal artificial intelligence (AI) model of "mixture-of-experts" type, offering a first overview of the architecture of the future Qwen4 family. This architectural redesign focuses on both the attention mechanism, residual connections, integration and optimization. Qwen3.8-Flash-Next has a main model of 125 billion parameters, complemented by 51 billion additional parameters, of which only 6 …
Coverage Details
Bias Distribution
- 75% of the sources are Center
Factuality
To view factuality data please Upgrade to Premium









