Qwen 3.6-27B just embarrassed a model 14X its size.
Alibaba built a 27B parameter model that beats its own 397B flagship on coding. But the size difference isn't even the most useful part. Why Qwen 3.6-27B matters: → 27B dense model — all parameters activate on every token
I also added a function to voice change only the vocals. Originally a Suno song.
Everything from songwriting to lip-syncing music video production can now be completed locally. Workflow to connect and run MiniMax H3, LTX 2.5, ACE-Step, Qwen 3.8 27B from iPhone (CloseBox) TechnoEdgeJP
50% more context unlocked for Qwen 3.8 27b Q4_K_XL dflash 2 on a single RTX 4090 (24 GB VRAM)
I found a hidden VRAM tax in llama.cpp. By combining my custom 2 bit DFlash 2 drafter with one overlooked server flag, I just unlocked another +80,000 tokens of context. Qwen3.8-27B
You have reached the end of the archive
All of qwen38