Quad 3090s Hitting 90 tok/s + on Qwen 3.8 Flash on llama.cp | Definitely wouldn't mind running this…
Quad 3090s Hitting 90 tok/s + on Qwen 3.8 Flash on llama.cp | Definitely wouldn't mind running this model but I just feel like my current setup is the smarter one.
This is the cutting edge period!
Omarchy first Then any other agent harness of your choice. My rank? Hermes Leave openclaw alone they are not serious download qwen 3.8 8b flash model and run locally
Anthropic said in the early morning that Fable 5.1 saves 45% on long tasks, and Artificial Analysis…
Anthropic said in the early morning that Fable 5.1 saves 45% on long tasks, and Artificial Analysis said in the middle of the night that each task is 20% more expensive. Neither side lied. It is smarter and can eat more tokens. It saves your time, not your bills. On the same day, Qwen 3.8 Max reached the top of Code Arena, with official $2 entry and $6 exit.
Qwen 3.8 27B
Everyone argues Q4 against Q5. The thinking budget moves quality seven times more than the quant does. Same weights. Same tasks. Effort off: 61.3%. Effort max: 80.3%. Twenty points. No quant in the same study moved it more than three.
You have reached the end of the archive
All of qwen38