Experiment with Qwen3.8-Flash-Next 176B Model on NVIDIA GB300 NVL72 for Agentic Coding
Alibaba releases weights for the Qwen3.8-Flash-Next 176B model, a preview of the upcoming Qwen4 architecture. This multimodal mixture-of-experts model activates 6B parameters per token and supports a native 262,144-token context window. Developers can experiment with these weights for agentic coding tasks on NVIDIA hardware.