Ran the "Lord of the Rings intro" test with both DeepSeek-V4-Flash 0731 (cloud version) and Qwen3.8-27B (local on my DGX Spark)
DS4 was quicker (40 min vs. 2hr 17 min) and hit some of the elements better (party tent), but the Qwen model made a much visually richer world and did
Me when I use Deepseek after my Claude usage runs out
Stanford just dropped a full course on self-improving AI agents.
The professors built Claude and Gemini. "we sampled the model 10,000 times. only 3 answers were correct. but those 3 solved what the best models in the world couldn't." 0:00 – scaling laws: why bigger models keep
Deepseek-v4-flash-abliterated seems to be such a chill guy afterall.
Is this what Anthropic is fearmongering 🤣 Told it to generate one short video using minimax h3 and a short music sample using minimax music3 and use its own taste, then connect them together. To me, it
You have reached the end of the archive
All of deepseek