FRAMEWIREIndonesiaUpdated Aug 16Live wire
0:00 / 0:00

Qwen 3.8-27B (Golden Shore) vs Opus 4.6 (Coastal World)

Same exact prompt, both were given only one shot to see what they would make.

DanielAug 15
0:00 / 0:00

Update on the free community Qwen 3.8-27B endpoint (was getting a bit too slow)

2x H200 replicas (+1 replica) now use speculative decoding (70 → 126 tok/s, measured) default thinking is now medium when not set (xhigh burns an enormous amount of tokens)

Victor MAug 1515
0:00 / 0:00

Been using Qwen 3.8 27B (Q4) locally on 64GB of VRAM.

Here is the verdict: SLOW 18 tps with ZERO system prompt to process and that degrades significantly with a harness system prompt and as the context window grows. RIP if you have to compact. I had it implement this PRD

Burke HollandAug 1534
0:00 / 0:00

Absolutely cooked with their models.

Since I'm pretty lame I've asked to install the Qwen 3.8-27B for me on my 1080... It got installed and the speed was ~1.8 tok./s. I've then decided to ask Daybreak Blue High to find ways to improve the speed to something more

SimonasAug 15