Google just mogged this tame impala test.
🌀 I gave 4 fast models the same Currents cover. Same prompt. One shot via aimlapi. • Gemini 3.8 Flash 2:58 • DeepSeek V4.1 Flash 2:54 • Muse Spark 1.3 3:34 • Qwen 3.8 Omni-Flash 3:50 Four completely different interpretations.
Constantly amazed at what you can do at home with *significant* investment (but also not boat money).
My own agent/harness running entirely on my own hardware. It's the Qwen 3.8 Flash model running on one Spark, plus all the voice models running on a 3090. It's plugged into OMP,
Testing out Jev. I am putting Jev against Qwen 3.8 27B in a chess game.
So far I keep getting no winners. But Jev's speed is uncomparable.
I have to retract what I posted as the fastest speed of Qwen 3.8 Flash on a single DGX Spark.
Cruz shipped a custom ExLlamaV3 fork and a native engine path. I ran the same model on the same Spark: 102.6 max, 80 tok/s on code, 70 tok/s on prose. My previous benchmark and
You have reached the end of the archive
All of qwen38