MTP nearly doubled Qwen 3.8 27B on my 4090.
It also filled the card to the last 112 MB. Qwen 3.8 dropped today with a draft layer baked into the weights. The idea is simple: a small attached head guesses the next couple of tokens, the main model checks all the guesses in one
My first prompt deepseek v4 flash inside deepseek harness preview.
This is crazy
The Silent Deepseek Moment - which will cause great damage to US AI companies.
We're giving you a $10 coupon to build on Ori Harness this weekend
What will you build? Run your favorite agent - Claude Code, Codex, DeepSeek, and more - with 500+ models, 70+ providers, all in one account, on OpenRouter Redeeming is easy:
You have reached the end of the archive
All of deepseek