Yume_X @yume_arasaki · 2,417 followers · 2d ago
Insane developments have been happening on Qwen 3.8 Flash. In 2 weeks, suddenly everyones computer can run Qwen 3.8 Flash.
I wanted to find out what happened and why everyone in local AI is using this model now.
Every number below is a receipt I verified, my own hardware or community, and the ones from my fleet live in my bench repo. Here is what I found.
WHY THIS MODEL
Qwen3.8-Flash-Next is a 180B parameter MoE that only activates around 6B per token.
On Artificial Analysis it scores 40 on the intelligence index against Claude Opus 4.8 at 42.
It beats Opus on Terminal-Bench 4.0 (25 percent against 22), AutomationBench (56 against 46), and GDPval. Opus still wins Humanity's Last Exam (49 against 38) and SWE-bench Pro (69.2 against 62.5, each vendor's own harness).
My read: Opus-class on agentic and coding work, one step behind on the hardest reasoning.
People call it Qwen 3.8 Flash. Same model.
WHY IT RUNS ON ALMOST ANYTHING EXPLAINED
Traditional LLM understanding is eve
392 38 27.2K+0 likes/h +13.8 views/h