FRAMEWIREIndonesiaUpdated Sep 3Live wire
0:00 / 0:00

My company's harness for Qwen 3.8-27B, Muse Glimmer 30B and Gemma 4:31B.

Local Laptop LLM

DanielSep 21
0:00 / 0:00

Bro… someone just built in one shot a Mario clone locally with Qwen 3.8 27b

Link:

Joman 👺Sep 21
0:00 / 0:00

Gemini 3.8 Flash High effort Vs Qwen 3.8 27B Max effort

BenSep 2
0:00 / 0:00

A 125B MoE model Just hit 25 tokens/sec on a single RTX 4090 at home.

Qwen 3.8 Flash Next + MTP speculative decoding at - 80k context. - 25.35 t/s decode - 471 t/s prefill Running a 125B Mixture-of-Experts model on a single consumer GPU

Md Ismail Šojal 🕷️Sep 22