Skip to the news

Versus

Alibaba, large language models

Qwen3.8-Flash-Next

Alibaba releases Qwen3.8-Flash-Next mixture-of-experts model achieving cost efficiency improvements.

Qwen3.8-Flash-Next Another open weights model from Qwen. This one is "a multimodal MoE model that also serves as an early preview of the architecture used in Qwen4". It's pretty big: 125B tokens, but only 6B active which means it gets a significant performance boost. I've been trying it out on a DGX Spark using these Unsloth quantized models. I'm still exploring the model - so far I've tried the 72.5GB UD-IQ1_S one (producing these pelicans) and the 78.9GB UD-Q2_K_XL (producing these). My favorite so far was this xhigh reasoning effort one from UD-Q2_K_XL: Via Hacker News

Simon Willison also at Hacker News, The Decoder

A quick introduction to Versus

Latest AI News

  • Ten stories per day, ranked by relevance.
  • Free, zero ads and no account required.
  • Designed the way we prefer to use it.

Choose sides

  • The teal square is biased towards robots.
  • The pink triangle gives you a combined view, unbiased towards either side.
  • The amber circle is biased towards humans.

Install it as an app

  • Quick and easy access.
  • Also works offline.
  • Receive alerts.

Ready for more?

  • We've got an RSS feed. (Yes, we see the irony.)
  • Subscribe to our newsletter for daily digests.
  • Buy merch and show whose side you're on.

Spread the word

PWA successfully installed.
Take that, Apple! 😉