FRAMEWIREIndonesiaUpdated Aug 17Live wire
0:00 / 0:00

We put Qwen 3.8 27B on a stock office mini-pc with 32GB of memory with no GPU, and let it rip.

Nothing crazy but super dope what we can get done at the edge.

Michael Westbrooks IIAug 171
0:00 / 0:00

Qwen-3.8-27B is used in Claudecode of M2Max. Since the memory bandwidth is only 400GB/s

Qwen-3.8-27B is used in Claudecode of M2Max. Since the memory bandwidth is only 400GB/s, after unremitting efforts, it can only reach the speed of 16token/s for the time being. With cache, the local machine is still too slow. It takes half a day to process 20,000 tokens at a time.

lifccAug 17
0:00 / 0:00

Qwen 3.8-27B just matched a model that was the best in the world 6 months ago.

And it runs on your laptop. 27 billion parameters. Tiny by AI standards. It keeps pace with Claude Opus 4.6 on coding, agent work, and reading images. Here's the crazy part: → Apache 2.0 license.

Julian Goldie SEOAug 17
0:00 / 0:00

If you're curious what ~48 tok/s looks like.

Running MTPLX Qwen 3.8 4 bit quant -- fully locally on 128gb MBP M5 from

Zach WillsAug 17