I did the 0.1.0 without the gauntlet loop and the 0.2.0 with the gauntlet loop
0.1.0 did all the hard work ( 12% weekly usage ) and 0.2.0 just did gauntlet loop graphics remaster spin loop ( 25% usage ). It got some results but it seems really token hungry.
I'll be working 8-5 starting tomorrow, I'm starting my job, don't go crazy without me, calm down
Benchmarks don't show a model's true capabilities, but specific tasks do.
The same task for GPT-6 Astra, Fable 5.1, Grok 4.6, and Opus 5 A tank game with voxel graphics, for 2 players via WebSocket and three.js
GPT-6 Astra is truly terrifying 😨
As a designer, I was confident that artificial intelligence would not outperform me in design... Until Astra came and turned everything around 🔄 He actually uses Figma, builds entire pages, from BI to UI with very high quality. Fable 5, Opus 5, even GPT-5.6 haven't reached this level. The question that concerns me…
You have reached the end of the archive
All of Claude Opus 5