Today we’re launching Portable Computer on NVIDIA DGX Spark.
Portable Computer is a fully local version of Perplexity Computer, where the entire runtime: orchestrator LLM, subagent LLM, agent harness all run on your local hardware. No cloud dependency.
Tried running the Qwen 3.8 2-bit model locally on my Book4 Pro running Omarchy✌️
It works, but I got around 1.2 tkn/sec ... way too slow to actually use it (expected).. this is pretty much the limit of what I can run on my current hardware... might just wait for 35B MoE model
Calculate the recent popular models and hardware data: M5 Max Qwen 3.8 27B MLX 4bit (dense model)
Basic short contextPP about 900-925 tok/sTPS 32-33 tok/s ~32k TPS and 28-29 tok/s In most cases, there is a significant improvement after turning on MTP or DFlash. The actual measurement of MTPLX is mostly 55-65 tok/sDFlash2, some people have reached about 70 tok/s M5 Ultra Today…
Warcraft III menu round 2
Gemini 3.1 Pro DSV4 Pro GLM 5.3 Grok 4.6 Fable Sol Kimi K3 I have opinions on which did the best. I think it's worth doing a 4 square comparison with Qwen 3.8 as well.
You have reached the end of the archive
All of qwen38