FRAMEWIREIndonesiaUpdated Aug 15Live wire
0:00 / 0:00

Qwen 3.8 is a game-changer!

It's performing on par with the latest Opus model. Seriously impressed with the capabilities here.

Bryan DowningAug 15
0:00 / 0:00

Finally tested Alibaba_Qwen Qwen 3.8 27B on two RTX 5090 GPUs.

🔥 Getting around 120–130 tokens/sec with vLLM, UnslothAI NVFP4, native MTP, and a 150K context target. The coding output looks very good.

RajuGangitlaAug 155
0:00 / 0:00

After deploying Qwen 3.8-27B locally, I learned another knowledge

Dgx spark is not suitable for running Dense models Different model architectures have different throughput requirements for underlying computing power and memory bandwidth. If a dense model does not do quantification and KV management well, no matter how powerful the hardware is, it will not work. Sure enough, practice brings true knowledge, let’s feel the speed.

LonelyAug 1515
0:00 / 0:00

QWEN 3.8:27b made this as a one shot prompt for me in 20m!

All while running locally on my machine. Kinda crazy

PawelAug 15