DeepSeek-V4-Pro-0813 is live on CoreWeave Serverless Inference.
1.6T parameters. 1M context. Built for long-horizon agent work, and priced for it: cache reads run $0.044/M for all the context your agent re-sends every step. Start here:
DeepSeek put out a vision version of V4 Flash and the repo weighs 305B parameters.
The text-only one is 304B. So vision cost them a billion parameters, which would be remarkable if that were the whole story. A month ago I went through the text repo. The model itself is 284B. The
GLM 5.3, GPT 5.6 Luna, Claude Opus 5, Grok 4.6, DeepSeek V4, Kimi K3, and a cloud browser for agents - all FREE.
DuckDuckGo: GPT 5.6 Luna and older models, no signup: LM Arena: free side-by-side, Opus 5, GPT-5.6 Sol, Grok 4.6, Qwen 3.8 Max:
Big tech doesn't want you to know this, but you don't need to pay $20/month to have ChatGPT with web search and document analysis.
I discovered Open-WebUI, the definitive open-source interface to run LLMs on your own machine (with total privacy and 0
You have reached the end of the archive
All of deepseek