Qwen3.8 Flash-Next as a Labs experiment, open for a short test window: temporary, not a permanent id. Served from Qwen's official FP8 checkpoint: a 125B mixture-of-experts model with 6B active parameters per token, the first open release of the next-generation architecture that powers the upcoming Qwen4 family, with native image and video understanding, built for fast coding and agentic tasks on a 256K context window. It thinks by default at xhigh effort; reasoning can be tuned down (low, medium) or turned off (none). Access is seat-gated through the Labs page while an experiment is live. It is offered at limited capacity and low availability, so expect it to be flaky and to go down under load: crash it, give it a moment, and try again. For production work we recommend umans-coder or umans-deepseek-v4-pro-0813.