Qwen 3.8 max vs deepseek v4 flash 0731 vs kimi k3 vs gpt 5.6 sol – on rubik's cube and chess
Four frontier models built a rubik's cube stand and solved it, then built a chess board and played claude opus 5 on it the setup: nousresearch's…
Four frontier models built a rubik's cube stand and solved it, then built a chess board and played claude opus 5 on it the setup: nousresearch's…
What top models like Claude Opus 5, Kimi K3, GPT-5.6, Claude Fable 5, Qwen 3.8 Max can't achieve.
Kimi k3 - fable 5 - opus 5 - gpt 5.6 with all the test results it’s evident that qwen 3.8 is the best chinese model currently live qwen 3.8 beats…
We ran all three on two one-shot 3D hero-object prompts 1)Fabergé egg on a black background 2)vintage diving helmet on a black background results:…
GPT-6(Astra)、Fable 5.1/5.5、Grok 4.6、Gemini 3.5 Pro、GLM-5.3、Qwen 3.8、Kimi K3.1… 开源吊打闭源,价格还只是零头 中国AI,8月掀桌子,一起期待!🇨🇳
And lowkey these creative designs ate They all understood the assignment, even tho Kimi K3 kept glitching and being buggy like 4-5 times 😂
Qwen 3.8 Max. 2.4 trillion parameters. 10+ days of autonomous coding. 365-day strategy planning.
Qwen 3.8 Max, Claude Opus 5, Kimi K3, and GPT-5.6 Sol were all given the same photo of clouds and asked to draw an animal silhouette based on their…
这次又准备对抗Workbuddy? 就从肉眼来看,技能Workbuddy确实内置的多一点。 实用性没有实操,不敢乱说; 但是千问这个IM频道还是很实用的,接着飞书。确实能干很多事情。
Only these effort levels did them correctly Interactive 3D ball-balancing platform using Three.js
Not just another "bigger AI model." This one activates only 95B of its 2.4 trillion parameters, yet ranks among the world's best frontier models.
While the new DeepSeek V4 Flash struggled on this benchmark. Which models should I compare next?
It is starting to make closed models look overpriced Alibaba just dropped Qwen 3.8 Max: a 2.4T parameter open-weight model coming out of China, in…