FRAMEWIREIndonesiaUpdated Aug 16Live wire
0:00 / 0:00

Been using Qwen 3.8 27B (Q4) locally on 64GB of VRAM.

Here is the verdict: SLOW 18 tps with ZERO system prompt to process and that degrades significantly with a harness system prompt and as the context window grows. RIP if you have to compact. I had it implement this PRD

Burke HollandAug 1534
0:00 / 0:00

Absolutely cooked with their models.

Since I'm pretty lame I've asked to install the Qwen 3.8-27B for me on my 1080... It got installed and the speed was ~1.8 tok./s. I've then decided to ask Daybreak Blue High to find ways to improve the speed to something more

SimonasAug 15
0:00 / 0:00

How to run Qwen 3.8 27b (Opus 4.6) At Home 🏆

Pushing 30+ tk/s on Mac M2 Max // 97 tk/s on 4090 TL;DR: Using Codex + GPT‑5.6 SOL at xHigh reasoning, I stress-tested the new Qwen3.8‑27B across my local hardware and nearly doubled its speed through native speculative decoding.

Eric ⚡️ Building...Aug 1514
0:00 / 0:00

Qwen3.8-27B fits in 24GB.

The agent that scored 73.0 on Terminal-Bench does not. Qwen dropped 3.8-27B yesterday under Apache 2.0 and the number everyone is repeating is the community Q4 GGUF at roughly 17GB, which makes a 24GB card sound like enough. I have spent four days

BountyAug 1514