I ran Qwen 3.8 27B on a single 8GB GPU.
IQ4_XS + Q4 KV cache at 70k context. The token speed hovers around 4-5 t/s. Slow but it's still amazing that you could get a result that can outperform even Claude Opus 4.6 on some of my test. All in one-shot with zero agent looping.
To deploy Qwen 3.8 27B locally, what kind of hardware configuration is required and what are the options?
Modeling challenge! 0x Alpha vs Grok 4.6 vs Qwen 3.8 big showdown!
Token Usage + Cost + Highlights Full Disassembly👇 📌 Core Data • Qwen 3.8: About 550,000 tokens, $1.90 • Grok 4.6: ~2-4 million tokens, $4.20 • 0x Alpha: ~1.2 million tokens, $0 (free) 📌 Comments on the highlights of each model 🔥 Qwen
Qwen 3.8 Max Full COURSE 1 HOUR (Build & Automate Anything)
You have reached the end of the archive
All of qwen38