Contrary to the previous example
Contrary to the previous example, this is an example of a failed prototype where I tried using Jev but it was not effective and ended up costing me more. If you combine Jev and Groq, wouldn't it be super fast? I tried that, but as a result, Groq alone was sufficient, but the cost jumped 20 times...sweat…
Here’s my first test of Qwen 3.8 27B on the inco_ai Splash inference engine.
✅ More than doubled the tokens/s that the M5 Pro was getting on oMLX. ✅ Two command setup ✅ Equal Game/Code quality
Bottom line! It’s hard to build a company and have big IPO’s when open weight models provide IP security and a less expensive solution.
Grok rank Qwen 3.8 vs common cloud models.
One AI model now watches and listens to your whole video
Qwen 3.8 Omni Flash turns it into a recap without chopping anything up.
You have reached the end of the archive
All of qwen38