DeepSeek v4.1 Flash running on 2x DGX Sparks
2× dgx spark owners rejoice!
DeepSeek-V4.1-Flash at 3.30 bpw EXL3, targeting just TWO DGX Sparks. 🔥 And yes, VISION is included. DeepSeek already ships its routed experts in FP4. I pushed that expert bank to a 3.30 bpw using my internal SAGE-EXL3 dynamic quant tool average
SemiAnalysis founder Dylan Patel says DeepSeek's models run worse on Google's TPUs and Amazon's…
SemiAnalysis founder Dylan Patel says DeepSeek's models run worse on Google's TPUs and Amazon's Trainium, and DeepSeek does not care about any chip that is not Nvidia or Huawei [ What are some of the ways that we've either designed our models for our hardware or even ways that…
We have released v0.26.14 of AITuber OnAir Core, expanding the options for LLM/TTS in AITuber development!
✅️Support for multiple LLM models such as GPT-6 Astra, Claude Fable 5.1, DeepSeek V4.1 Flash, OpenRouter models ✅️New support for Inworld TTS-2 Flash in TTS
You have reached the end of the archive
All of deepseek