Qwen 3.8 Max has arrived — but does it really deliver all of this?
👀 In today's video, I reacted to the SWEN benchmark and commented on the performance of the new model, comparing the results and raising an important question: to what extent do the benchmarks really show the ability to
Two Qwen 3.8 27B agents, a designer and a coder, went back and forth more than 50 times on one desktop before calling the project done.
128GB of memory, nothing metered. That round count is what got me. You'd never let a loop run that long against a paid API. You'd cap it at
Testing Qwen 3.8 Flash Next Q4 on the DGX Spark: Day 34.
Asked it to create Pac-Man, and after 24 minutes, it failed. DeepSeek is the only model that's come close. Interesting to see how future versions progress!
WTF guys, why is literally nobody yapping about qwen 3.8 flash rn?!
Tested the exact same one-shot flappy bird prompt, and the speed-cost-output balance is actually insane solid 8/10. yes some barrier placement is kinda weird, and GLM-5.3 flash has slightly smoother visuals,
You have reached the end of the archive
All of qwen38