原帖内容
Qwen 3.8 Flash Next (125B A6B MoE) Vs Qwen 3.8 27B (dense) - Both Q4_K_XL - Reasoning Off Prompt: Create a very high quality realistic video like animation of the solar system. three.js via cdn. Let the camera revolve around. single html file. dont call any tools. dont need any controls. Qwen 3.8 27b focused on computational physics & geometry (inclinations, coordinate math, moon orbits). While Qwen 3.8 125B A6B focused on cinematography & visual UX (lighting contrast, atmospheric glow, framing, and HUD overlays). It also followed the camera instruction correctly and produced result thats visually more stunning. Both the models run on single RTX 3090/4090 (24 GB VRAM). the 27b will fit entirely in the VRAM, but you'd need 100-120GB RAM for the new 125B MoE (check out the previous post for complete inference benchmarks and llama.cpp setup) Qwen 3.8 Flash Next beats Qwen 3.8 27B in almost all the benchmarks. have dropped the intelligence benchmark comparison and unsloth quant huggingface links in the replies. Which one would u be running regularly on your hardware?








