FRAMEWIREIndonesiaUpdated Aug 25Live wire
0:00 / 0:00

Atomic Chat HQ rented a B200 cluster

Fired up Alibaba's 2.4-trillion parameter Qwen 3.8-Max model, and told it to build a full Call of Duty clone from a single prompt. The media thought that was the story. It was not. The crew thought that was the story. It was not. The story

Pascual ⚡Aug 24
0:00 / 0:00

Four frontier models took mazebench today and all four scored 0%.

The one that beat them scored 1% The account posting this built the benchmark, so it is the source and not a screenshot of a screenshot. Ox Alpha zero. Grok 4.6 zero. GLM 5.3 zero. Qwen 3.8 Max zero. Gemini 3.7

StarHazeAug 243
0:00 / 0:00

Every benchmark you have ever quoted was run in the cloud.

This one ships 10,000 results off actual phones and laptops Liquid AI and Artificial Analysis released Pipette today, and the interesting part is not the tool, it is what it admits. Cloud benchmarks measure a model.

StarHazeAug 247
0:00 / 0:00

The Qwen 3.8 27B Alibaba_Qwen is clearly a monster no other model comes close

Even GLM-5.2 failed to produce such a polished simulation in a single shot without needing follow-ups to point out missing elements. #qwen #qwen38 It's so impressive to obtain that on a 16g vram gpu

DamienAug 24