ZAI pulled off the best stealth launch of the year
Last week they anonymously released glm 5.3 flash as oxalpha and it became the most popular model of the week today they finally revealed it and the biggest highlights are: cheapest model in the glm 5 family around 10x
Deepseek V4 Flash 0731 TP=4 Single-Stream Decode
420 tok/s Community record with prefill off was 357.9, and I was able to reach 446.4 tok/s due to two variables + 6000 memory overclock. Recipe and receipt below as always.
LongHorizon-Harness research showed benchmark task completion jump from 51.8% to 80.7% using one architectural shift
Separating coding harness execution from a verified, audited task state. in this 3-minute breakdown, Claude Code demonstrates direct terminal execution and file
Q&A on Agentic Operating Systems.
The best questions from the community. Team access to your agent OS? Skill plus Tailscale. Everyone connects to one machine. Which model for what? Claude Fable 5 to orchestrate. DeepSeek V4 Pro for agentic work. GLM 5.3 for coding. Want it on
You have reached the end of the archive
All of deepseek