Testing out Jev. I am putting Jev against Qwen 3.8 27B in a chess game.
So far I keep getting no winners. But Jev's speed is uncomparable.
I have to retract what I posted as the fastest speed of Qwen 3.8 Flash on a single DGX Spark.
Cruz shipped a custom ExLlamaV3 fork and a native engine path. I ran the same model on the same Spark: 102.6 max, 80 tok/s on code, 70 tok/s on prose. My previous benchmark and
There's been a lot of hype about Ternary Bonsai2 but fw demonstrations on how it actually compares to the full 16 bit version of Qwen 3.8 27B.
Since I have my own quantized version of that model thatI'm testing, ~12gb file size, I decided to throw Bonsai2 in there also to get a
Now I tested the difference between Qwen 3.8 27B local vs GLM 5.3 both on Droid
There is a difference in quality and being able to follow the images and attention to detail more closely. Clearly GLM 5.3 is ahead in quality, although it took much longer to build the game.
You have reached the end of the archive
All of qwen38