原帖内容

Hy4 preview just dropped, so I tested it on WorkBuddy against three other models: DeepSeek V4 Flash, GPT-5.6-Luna, and GLM-5.3... the top 4 models on this week’s OpenRouter coding leaderboard. The best part? They’re all available in WorkBuddy, so you don’t need to download multiple Agent apps. For the test, I used one single prompt to build one of those games you’ve definitely seen in ads: swipe/move through number gates, make calculations, and eventually defeat the boss. The results were pretty interesting 👇 For a task of this difficulty, everyone performed well except DeepSeek. It seemed to finish coding without checking for bugs, and the game couldn’t run properly. GPT-5.6-Luna worked, but felt like an arithmetic multiple-choice quiz. A little boring. GLM-5.3 and Hy4 preview were much more impressive. GLM was slower (video sped up 2.2×) and only had three movement positions, while Hy4 offered continuous horizontal movement and a more natural speed. What surprised me most was Hy4’s game design: besides basic arithmetic, it added weapons that reduce enemy damage or increase stolen power, making the game much more fun. Also worth noting: Tencent’s Sherry quantization compresses the 770B MoE model from ~1.5TB BF16 to 214GB, while keeping performance close to the original... making local deployment much more practical. Hy4 preview is free on WorkBuddy until September 10. Give it a try 👇 https://www.workbuddy.ai/