Qwen-3.8-27B NVFP4 (weight 21.3G), MTP is not enabled, the throughput speed on DGX Spark is >10tok/s
Feel the speed👇🏻 I will turn on MTP later for comparison.
Gemini 3.7 Flash test against Qwen 3.8 and Grok 4.6 and GLM 5.3;
Same recipe, 4 models: a harvester and farmers working. to be honest Qwen 3.8 did exactly what was asked of it. Grok 4.6 failed to create textures. Which one do you think worked better?
What a crazy week in AI!
🚀 DeepSeek V4 0813 GLM 5.3 Grok 4.6 Qwen 3.8 Max Qwen 3.8 27B Index TTS 2.5 MiDasheng LTX 2.5 GPT Ultrafast Gemini 3.7 Flash MAGI-2 WorldClaw Nemotron Lightning & more! Watch the full recap:
Okay I'm officially blown away and SUPER HAPPY with qwen 3.8!
I did all the model testing and running a Q16 MTP model that is ripping fast now Overclocked the system and GPU to handle the dense memoy calls better Generated my first game and I cannot believe how good it looks
You have reached the end of the archive
All of qwen38