Qwen 3.8 27B is on another level of Game Creation.
This wasn't a one-shot prompt. It was a two-shot prompt to get to this game and I had to find a bug that prevented the game from showing on screen. But its all HTML / Javascript. It built a Retro Invaders clone with a
How does an LLM inference engine work under the hood?
I wrote an article on my Qwen 3.5/3.8 engine for Apple GPUs: KV/prefix caches, DFlash/MTP speculative decoding, and deferred linear attention state commits. It achieves a 1.25× speedup over llama.cpp.
I have been working with cerebras inference running qwen 3.8 27b and the first issue I see is reviewing…
I have been working with cerebras inference running qwen 3.8 27b and the first issue I see is reviewing the wall of output this thing puts out in seconds.
Many big guys are asking today, can one of your computing cabins run Qwen 3.8 Flash Next + Blender?
Of course you can, but my modeling skills are relatively poor, so the build is ugly 😂 If this AI solution is handed over to professional design engineers, it should be very beautiful.
You have reached the end of the archive
All of qwen38