FRAMEWIREIndonesiaUpdated Sep 10Live wire
0:00 / 0:00

35B MoE. 8GB of VRAM.

39.3 tokens/s decode. That is Qwen3.6-35B-A3B NVFP4 on a laptop RTX 4060 — not an H100 rack. The hot experts sit in a GPU cache. The rest of the pool lives in system RAM. The paper’s coding-agent run even clears the 33 tok/s Codex median they cite. 🆕

EchoGitSep 10
0:00 / 0:00

Horse tinder one shot Deepseek V4.1 Flash.

I told it to make a video too and it did!

Vstalin GradySep 10
0:00 / 0:00

Every time a new model such as DeepSeek V4.1 Flash comes out, founder Liang Wenfeng seems to be forced to do this dance for the time being.

西山雄大「やさしい文書」tAIoSep 10
0:00 / 0:00

DeepSeek V4.1 just launched, and here’s how to use DeepSeek models for completely FREE.

No subscription. no API credits. no card required. 1️⃣ open up VS Code 2️⃣ go to Extensions, search for Cline, then install the extension if you don’t already have it. 3️⃣ open Cline,

m0hSep 1015