Every benchmark you have ever quoted was run in the cloud.
This one ships 10,000 results off actual phones and laptops Liquid AI and Artificial Analysis released Pipette today, and the interesting part is not the tool, it is what it admits. Cloud benchmarks measure a model.
The Qwen 3.8 27B Alibaba_Qwen is clearly a monster no other model comes close
Even GLM-5.2 failed to produce such a polished simulation in a single shot without needing follow-ups to point out missing elements. #qwen #qwen38 It's so impressive to obtain that on a 16g vram gpu
I plugged Qwen 3.8 27b into DeepSeek Harness and asked for a replica Age of Empires 2
Returned to my Studio 70 minutes later to find the Mac still purring... So far my local model tests on this prompt have produced unsharable results Let's see what we get
Qwen 3.8 27B Q4 quant destroyed Q8 at voxel island creation ⛏️
You have reached the end of the archive
All of qwen38