Qwen3.8 Flash Next is the MoE/new architecture they released instead of a qwen3.8 35b MoE. They didn’t say it’s “better” they just mysteriously said 35B was not what they were releasing this round. A few days later, they released the flash next model weights.
EDIT: and it’s great, doesn’t overthink, but slow AF on my Poweredge T640 server with PCIe3 serving my RTX3090 and DDR4 (even in a proper quad channel config). Best I can get is 15 T/s
Qwen3.8 Flash Next is the MoE/new architecture they released instead of a qwen3.8 35b MoE. They didn’t say it’s “better” they just mysteriously said 35B was not what they were releasing this round. A few days later, they released the flash next model weights.
EDIT: and it’s great, doesn’t overthink, but slow AF on my Poweredge T640 server with PCIe3 serving my RTX3090 and DDR4 (even in a proper quad channel config). Best I can get is 15 T/s