Here's Qwen 3.8 27b in action using the DeepSeek harness.
This is self-hosted from our facility. 55-65 tokens per second. Text go woosh. This is FP16 precision 262k context vision enabled. We will be adding more open weight models soon, next up will be Qwen 3.8 Flash Next
OpenAI charges ~$50 per million tokens.
DeepSeek can do it for as low as 60 cents. At the All In Summit, friedberg asks $MSFT CEO satyanadella directly: do frontier labs have the wrong business model?
Today's update: 1. Guardrails are live.
Set policies that check every request and response, and block or redact what you don't want through. 2. Server tools are live: minirouter:web_fetch and minirouter:datetime. 3. Send response_format with a JSON
My first Multiplayer Game 100% vibe coded with Astra, Fable 5.1 and Deepseek 4.1F is out.
Check it out, and come play! :) I can't begin to express how THRILLED I am. This took maybe 7-8h in total (with one of the most advanced custom harness ever) :-)
You have reached the end of the archive
All of deepseek