Just shipped the best model you can run on a single DGX Spark and the results are impressive.
Qwen 3.8 Flash-Next, NVFP4, one GB10, 128GB. I measured it today. I'm comparing 1 DGX Spark vs 2 DGX Spark on this model so you don't have to. 27.3 tok/s single-stream
Hermes + qwen 3.8 in telegram is the best combo
Testing out Qwen 3.8 27B Q3_K_XL ~20tok/sec on a 2060 laptop + 3060 rig connected via rpc
96k context #AIart️️️️️️️️️️️️️️️️️️️️️️️️️️️️️️️️️️️️️️️️️️️️️️️️️️️
There's nothing more infuriating than GPT-5.6 or Fable acting "morally superior" and refusing to do the work I assign them.
And it happens often... GPT-5.6 refusing: - To create content the way I say. - To create ad campaigns the way I want. - To advise or answer some
You have reached the end of the archive
All of qwen38