FRAMEWIREIndonesiaUpdated Aug 21Live wire
0:00 / 0:00

Open-weight Qwen 3.8 2.4T built a Call of Duty clone in one prompt 🪖

Alibaba_Qwen released the 2.4T Max weights, so we rented a B200 cluster and asked the model to make a Call of Duty clone Output: ~1.1M tokens · 5 hours · one prompt Almost no one can run 2.4T at home,

atomic.chatAug 2110
0:00 / 0:00

I had some fun playing around with the Qwen 3.8 model yesterday.

What blows my mind is that it's nearly a frontier model yet it runs on my 5 year M1 Max laptop! Now it's not super fast, 16 tokens/s but still! Check out the little video I made about how to set it up:

Matt RongeAug 211
0:00 / 0:00

I'm doing exactly this, my voice goes thru my home server running Qwen 3.8

daniel 🍳Aug 212
0:00 / 0:00

So I’m working on making decoding and prefill as fast as possible on my veloGB10 engine for the 3.8 27b model.

My 4x DGX Spark networking is switched (Mikrotik CRS812 DDQ ), High quality QFSP56 cables with all paths verified to run at max speed. This is a pure Rust/CUDA (with

Άντιpacman🪽Aug 21