[AI]■ STORY TIMELINE
ALIBABA'S QWEN3.8-FLASH-NEXT CUTS COSTS WITH NEW ARCHITECTURE
Alibaba's Qwen team unveiled Qwen3.8-Flash-Next, a mixture-of-experts model that activates only 6 of 125 billion parameters per token. The model achieves competitive performance at one-ninth the training cost of larger rivals.
Hacker News+0m
Article URL: https://qwen.ai/blog?id=qwen3.8-flash-next Comments URL: https://news.ycombinator.com/item?id=49448210 Poin…
The Decoder+1h 48m
Alibaba's Qwen team is previewing the Qwen4 architecture with Qwen3.8-Flash-Next, a mixture-of-experts model that activa…