Qwen3.8-Flash-Next
Qwen 团队发布 Qwen3.8-Flash-Next 开源模型,该模型拥有 1250 亿参数但仅激活 60 亿参数。作为多模态混合专家(MoE)模型,它被视为 Qwen4 架构的早期预览版本。用户已在 DGX Spark 上测试该模型的 Unsloth 量化版本,包括 72.5GB UD-IQ1_S 和 78.9GB UD-Q2_K_XL 版本,其中 UD-Q2_K_XL 在推理任务中表现最佳。
EVENT DOSSIER
The Qwen team released the Qwen3.8-Flash-Next open-source model on August 26, 2026. This model adopts a multimodal hybrid expert (MoE) architecture with a total of 125 billion parameters, but only 6 billion parameters are activated, making it an early preview version of the Qwen4 architecture. Users have tested the model’s Unsloth quantized version on the DGX Spark platform, including versions with 72.5GB UD-IQ1_S and 78.9GB UD-Q2_K_XL. The test results show that the UD-Q2_K_XL performs best in inference tasks.
Qwen Qwen3.8-Flash-Next 发布125B参数多模态MoE模型
Qwen 团队发布 Qwen3.8-Flash-Next 开源模型,该模型拥有 1250 亿参数但仅激活 60 亿参数。作为多模态混合专家(MoE)模型,它被视为 Qwen4 架构的早期预览版本。用户已在 DGX Spark 上测试该模型的 Unsloth 量化版本,包括 72.5GB UD-IQ1_S 和 78.9GB UD-Q2_K_XL 版本,其中 UD-Q2_K_XL 在推理任务中表现最佳。