FRAMEWIREEnglishDiperbarui Sep 17Kabel langsung
0:00 / 0:00

RTX 4090 saya baru saja mencapai 181 tok/s pada model Qwen 3.8 27B.

O Senin saya pikir 140 adalah batas atas. Saya salah. mr_r0b0t membuat mesin yang berbeda. Saya ingin mendorong kartu itu lagi untuk melihat apakah saya dapat memerasnya lebih banyak. Jumlah judulnya naik 30%, jumlah suka-untuk-sukanya pindah

Yume_XSep 1723
0:00 / 0:00

I ran Qwen 3.8 Flash Next on SSD and it's fast!

Using atomic_chat_hq AD-IQ4XS 85GB size Model weight to 24GB VRAM + 32GB RAM and the n-gram stream directly from SSD Result: 22 tok/sec (14 tok/sec @ 76k context) Output is still amazing since we're using 4-bit quant See

Red StaplerSep 171
0:00 / 0:00

No WiFi. No servers. No internet.

A builder just paired Qwen 3.8 27B with Cerebras chips to make an offline "browser" that doesn't load pages — it hallucinates them live, at 2,000 tokens/sec. Type any site. Pick any year. It invents a page that fits — 1999 web, 2045 web,

apikey_officialSep 174
0:00 / 0:00

My kid asked for a riddle monster to share with friends.

Here is Ridley! Tool Stack: Totem LLM: qwen 3.8 27B ComfyUI with ltx_io All locally running AI! Try it: Demo Site: You can make your own here:

Fred TerziSep 179