Benchmarks are starting to matter less.
GPT-6 Astra is hitting 99.9% on ARC-AGI-3 and 98% on FrontierMath Tier 4. Claude Opus 5 is SOTA on Frontier-Bench and GDPval-AA, while dominating AutomationBench. But Astra costs 2× more per output token: $50/M output vs $25/M for Opus.
The dispute is open: Opus 5 vs GPT-5.6 Sol.
🔥 They both received the same challenge. But which one delivered the best result? At KAIROGEN you have access to 30+ AI models.
Claude Opus 5 just keeps getting better
I had an idea in my head. Turned it into a prompt. And somehow… this came out.
Simple cad designs was solved earlier ..
This one was from Opus 5 .. this was built in about a day
You have reached the end of the archive
All of Claude Opus 5