This is the comparison that everyone is waiting for, four strong models facing each other 👀
1. The Qwen 3.8 is coming in strong 2. The GLM 5.3 is coming to take the position 3. Grok 4.6 has its own people who follow it 4. Gemini 3.7 It is not surprising that it is banned What are your expectations, who will emerge the strongest?
Ox Alpha is definitely NOT GLM 5.5, it’s honestly way worse than GLM 5.3
Just ran the benchmarks: Qwen 3.8 27B choked. Opus 5 still claps everything easily
Same prompt, four models.
Two are open-weight. Can you tell which? Gemini 3.7 Flash, Qwen 3.8, Grok 4.6, GLM 5.3. One test isn't a benchmark, but it shows why open models belong in the mix. Leaderboards get you a shortlist. Your own prompts tell you which one actually fits.
Qwen 3.6-27B just embarrassed a model 14X its size.
Alibaba built a 27B parameter model that beats its own 397B flagship on coding. But the size difference isn't even the most useful part. Why Qwen 3.6-27B matters: → 27B dense model — all parameters activate on every token
You have reached the end of the archive
All of qwen38