Upcoming deepseek v4.1 flash runs on ~300 tok/s
The preview endpoint still lacks vision, but I'm hoping the official release will be multimodal.
DeepSeek V4.1 Flash is a very interesting model.
I just tested it on the BridgeBench lava lamp test and it took LONGER to complete than Fable 5.1 and GPT 6 Astra. It ran at 344 toks/sec, spent 23.5M tokens with a cache hit rate of 99.7%, and cost $0.33. Even though it runs so
Reuters quoted two people familiar with the matter as saying that Chinese artificial intelligence…
First update of the day: GLM 5.3 Flash tops the value chart at $0.09
While DeepSeek V4 Flash offers the lowest price at $0.06.
You have reached the end of the archive
All of deepseek