# Qwen3.8-Flash-Next 发布：Qwen 开源多模态 MoE 模型，提前预览 Qwen4 架构

- 来源：Simon Willison 博客
- 发布时间：2026-08-27 07:52
- AIHOT 分数：65
- AIHOT 链接：https://aihot.news/items/cmtatjy570946roamigoxjc29
- 原文链接：https://simonwillison.net/2026/Aug/26/qwen38-flash-next

## AI 摘要

Qwen 发布开源多模态 MoE 模型 Qwen3.8-Flash-Next，作为 Qwen4 架构的早期预览。模型总参数 125B，仅激活 6B 参数，带来显著性能提升。作者已在 DGX Spark 上通过 Unsloth 量化版本试用，包括 72.5GB 的 UD-IQ1_S 和 78.9GB 的 UD-Q2_K_XL 版本。

## 正文

Another open weights model from Qwen. This one is "a multimodal MoE model that also serves as an early preview of the architecture used in Qwen4".

It's pretty big: 125B tokens, but only 6B active which means it gets a significant performance boost.

I've been trying it out on a DGX Spark using these Unsloth quantized models. I'm still exploring the model - so far I've tried the 72.5GB UD-IQ1_S one (producing these pelicans) and the 78.9GB UD-Q2_K_XL (producing these).

My favorite so far was this xhigh reasoning effort one from UD-Q2_K_XL:
