FRAMEWIREIndonesiaUpdated Aug 26Live wire
0:00 / 0:00

Alibaba's Qwen team just dropped Qwen 3.8 Max, a multimodal model with open weights.

It boasts 2.4 trillion parameters, a 1M token context window, and can create apps from screenshots. Best part? It's 80% cheaper than GPT-4.5 and 88% cheaper than Claude 3. | Peter Diamandis

LionblissAug 26
0:00 / 0:00

Qwen 3.8 Flash is now live in closed beta.

Join the waitlist:

Dime TerminalAug 264
0:00 / 0:00

The VRAM barrier is officially dead.

I just ran Qwen 3.8 Flash Next (MoE) 125B A6B with a 250,000 context window on a single 24GB RTX 4090. 21 tokens/sec decode. 364 t/s prefill. no mtp. no dflash. no kv cache quantization! We are running datacenter models on consumer

AlokAug 261.0K
0:00 / 0:00

Qwen 3.8 went to zero on two platforms today and one of them takes the word free as the API key

K2S is the only one in this whole free-model wave who put the base URL and the key next to the model names instead of a screenshot of a pricing page. The strongest thing in that post

StarHazeAug 26