FRAMEWIREIndonesiaUpdated Sep 22Live wire
0:00 / 0:00

Qwen 3.8 27b via Tesseract Server

What a phenomenal model that can be run on a Mac

SpokSep 22
0:00 / 0:00

Try out cache-to-cache here

ONE DGX Spark Working with MiaAI_lab recipe of Qwen 3.8 flash next, but you can use any model Simple test of an essay about the research paper:

droveSep 22
0:00 / 0:00

No single-stream headline today, so here's the part builders actually argue about

Two boxes running the same Gemma 4 12B IT QAT at 4-bit, and the spread between them is 3%. 29 tok/s on Strix Halo (Q4_K_XL). 28 tok/s on DGX Spark (Q4_K_M). Same model, different machines, and the

Zach ASep 22
0:00 / 0:00

Gained a lot of respect for Zuck lately.

Muse and Qwen 3.8 27B are now daily drivers. Seriously impressed with their latest AI developments.

StartupHakkSep 21