原帖内容
📢 @Alibaba_Qwen's Qwen3.8-2.4T-A95B weights just dropped, and SGLang already has day-0 support with DSpark. HF: http://huggingface.co/Qwen/Qwen3.8-2.4T-A95B Cookbook: http://docs.sglang.io/cookbook/autoregressive/Qwen/Qwen3.8 🖖 What kind of model can help optimize an inference engine? Qwen3.8-2.4T-A95B can take on complex engineering tasks, plan and iterate on its own, and keep going until the job is done. We even asked it to optimize its own serving on SGLang, and it ran unattended for 4.5 hours with verified performance improvements. Thanks to the Qwen team @Alibaba_Qwen, @NVIDIAAI, and @AIatAMD for the close collaboration.








