Advanced Frontier LLM Coding Benchmark 13.08.2026
Modeller: - Grok 4.6 Extra High - Cursor - Opus 5 Max - Claude Code (App) - Qwen 3.8 Max - Qwen-Code (CLI) - Kimi K3 Max - Kimi-Code (App) - GLM 5.2 Max - Zcode (App) - GPT 5.6 Sol Very High - Codex (App) -
I used grok with Grok Build to continue developing my hobby game in threejs almost no errors
Very fast but being fast also makes me prompt more meaning spend more as well, it is both good and bad in a way :) Almost no errors as well, flawless job. Also used for one my
New harness (qwen 3.8) got screenshots of the account.
Told it to write a song. it did. minimax-music3 made the track. first output.
Gemini 3.7 Flash/Qwen 3.8 Max by just being told "make it as realistic as possible"
You have reached the end of the archive
All of qwen38