digestweb.dev
Propose a News Source
Support usSponsor
🤝
Curated byFRSOURCE

digestweb.dev

Your essential dose of webdev and AI news, handpicked.

Advertisement

Want to reach web developers daily?

Advertise with us ↗

Back to Daily Feed

Qwen3.8-Flash-Next: A New Multimodal MoE Model Preview

Must Read

Originally published on Simon Willison's Weblog by Simon Willison

View Original Article
Share this article:
Qwen3.8-Flash-Next: A New Multimodal MoE Model Preview

Summary & Key Takeaways ​

  • Qwen3.8-Flash-Next is a new open-weights multimodal Mixture-of-Experts (MoE) model.
  • It serves as an early preview of the architecture planned for Qwen4.
  • The model is large (125B tokens) but has only 6B active parameters, boosting performance.
  • Simon Willison shares his initial experiences trying out quantized versions of the model.
  • This release is significant for the open-source AI community and model architecture research.

Our Commentary ​

A new open-weights multimodal MoE model from Qwen, and an early peek at Qwen4's architecture? That's big news for the AI community. The 125B total, 6B active parameter count is intriguing for performance. Simon Willison's early explorations are always a good sign of something interesting brewing.

View Original Article
Share this article:
RSS Atom JSON Feed
© 2026 digestweb.dev — brought to you by  FRSOURCE