FRAMEWIREIndonesiaUpdated Aug 14Live wire
0:00 / 0:00

Qwen 3.8 Max: A free open model just beat GPT and Claude at using a computer.

The test: click buttons, open apps, finish tasks on a real desktop. → Qwen 3.8 Max scored 86.1 → GPT 5.6 scored 83.2. Gemini scored 76.2. Claude scored 85.0. → Alibaba let it code alone for 16

Julian Goldie SEOAug 10
0:00 / 0:00

GLM 5.3 Max vs Qwen 3.8 Max vs Grok 4.6 vs Opus 5

GLM 5.3 Max is actually holding up really well here. My ranking for this run: Opus 5 > Qwen 3.8 Max > GLM 5.3 Max > Grok 4.6 Definitely competitive now. GLM is getting scary close.

OmedTheVibeCoderAug 141
0:00 / 0:00

All models are so creative now

I gave Gemini 3.7 Flash, Grok 4.6, Qwen 3.8 Max, and Claude Opus 5 an ace of clubs card. And told them to be creative and draw inside the card, using the card as inspo and here are their results

Ann NguyenAug 148
0:00 / 0:00

DeepSeek V4 Pro (0813) vs Qwen 3.8 Max

FoodTruck BenchAug 141