Comparison of Gemini 3.8 Flash with ChatGPT SOL, Opus 5 and Kimi K3
In this test, Gemini 3.8 Flash does not perform very well in the production of 3D content. This model provided such a weak output in about 2 minutes; Interestingly, in Goal Mode, the execution of the same task took about 30 minutes
Claude Opus 5 is INSANE
Look what I’ve just built.
We benchmarked the top models on our own coding tasks.
The results: - GPT 5.6 Sol (high) won on performance - Grok 4.6 (high) was the runner-up - GLM 5.3 Flash won on cost at comparable quality All 50%+ cheaper than our previous default (Opus 5)
The AM Brief, Tuesday September 1, 2026 Part 3
Good morning, It’s 8am in Miami and here is a recap of events that caught my eye. Part 3 Anthropic published landmark research demonstrating that autonomous Claude agent teams can independently drive AI alignment,
You have reached the end of the archive
All of Claude Opus 5