Anything under deepseek v4 flash basically doesn't exist to me.
With <200GB of vram, i'm getting ~800K context @ 200 tokens per second. It's not quite Cerebras fast, but damn. It is a joy to use, and its basically somewhere in the opus 4.7/4.8 range. It's incredibly good at
GLM 5.3 Max vs Qwen 3.8 Max vs Grok 4.6 vs Opus 5
GLM 5.3 Max is actually holding up really well here. My ranking for this run: Opus 5 > Qwen 3.8 Max > GLM 5.3 Max > Grok 4.6 Definitely competitive now. GLM is getting scary close.
All models are so creative now
I gave Gemini 3.7 Flash, Grok 4.6, Qwen 3.8 Max, and Claude Opus 5 an ace of clubs card. And told them to be creative and draw inside the card, using the card as inspo and here are their results
DeepSeek V4 Pro (0813) vs Qwen 3.8 Max
You have reached the end of the archive
All of qwen38