Qwen3.8-Flash-Next

2026-08-26 · 来源 Simon Willison

Qwen3.8-Flash-Next Another open weights model from Qwen. This one is "a multimodal MoE model that also serves as an early preview of the architecture used in Qwen4". It's pretty bi…

Qwen3.8-Flash-Next Another open weights model from Qwen.This one is "a multimodal MoE model that also serves as an early preview of the architecture used in Qwen4".It's pretty big: 125B tokens, but only 6B active which means it gets a significant performance boost.

I've been trying it out on a DGX Spark using these Unsloth quantized models.I'm still exploring the model - so far I've tried the 72.5GB UD-IQ1_S one (producing these pelicans) and the 78.9GB UD-Q2_K_XL (producing these).

My favorite so far was this xhigh reasoning effort one from UD-Q2_K_XL: Via Hacker News Tags: ai, generative-ai, llms, qwen, pelican-riding-a-bicycle, ai-in-china, nvidia-spark

TokenTria 策展

本文要点:英伟达、通义千问 新模型/版本发布。TokenTria 已收录新模型,可在「模型库」查看定价、上下文与能力标签对比。

全文转载自 Simon Willison,TokenTria 仅作信息聚合与策展引用。