FRAMEWIREIndonesiaUpdated Aug 19Live wire
0:00 / 0:00

Congratulations on the successful recovery of Zhuque 3! Today

ivey zenAug 19
0:00 / 0:00

I became a digital person. I didn’t expect it to be Uncle Liang. Listen to me, Uncle Liang, talk to me about deepseek.

Linx_OKX | 我爱MisaAug 192
0:00 / 0:00

DFlash 2 just dropped, and local inference is now on another level.

According to AA, Qwen3.8 27B sits roughly between GLM-5.2 Max and DeepSeek Flash 0731 Max. Q4_K_M running on a single 32GB RTX 5090: 75 tok/s sustained, up to 200 tok/s peak, 131K context, one active session.

Andrey BaksalyarAug 19
0:00 / 0:00

Qwen 3.8 27B Q4_K_M - 90 tokens/sec on a single NVIDIA RTX 4090 (24 GB VRAM) with Dflash2!

(MTP 60 tps -> 90 tps Dflash2!!!!) Local AI moves so fast (literally!) it’s terrifying. Z lab just dropped DFlash 2 for Qwen 3.8 27b and Muse Glimmer. I patched llama.cpp (PR #27342) and

AlokAug 1924