FRAMEWIREIndonesiaUpdated Aug 29Live wire
0:00 / 0:00

Qwen 3.8 27B thinking effort comparison, personally I feel the medium and thinking off not worth it, either go with thinking low or xhigh..

Note: quant may impact the model perf, looks like 4bit quant is way easier to get itself in thinking loop

mzbaAug 291
0:00 / 0:00

Tencent's Hy4 Preview just leaked before anyone was ready, and this breakdown covers exactly what's…

Tencent's Hy4 Preview just leaked before anyone was ready, and this breakdown covers exactly what's inside it, including the benchmark numbers most coverage is going to skip past. This video breaks down TencentHunyuan newly released Hy4 Preview, a 770 billion parameter mixture…

Lomash KumarAug 292
0:00 / 0:00

Day 35 of testing DGX Spark with DeepSeek Harness & Qwen 3.8.

Encountered setup issues, power loss, and slow token speeds (10-12 tokens/sec vs. expected 60). Memory usage is high (90GB). Long processing times (88 mins) suggest potential issues. Will revisit tomorrow.

Raymundo OjedaAug 29
0:00 / 0:00

Alibaba just turned AI from a chatbot into a workforce.

And the biggest upgrade isn’t the 2.4 trillion parameters. It’s what happens after you give it a job. What Qwenwork can actually do: → Research a topic, write the document, generate images, build the webpage, create

Julian Goldie SEOAug 291