Google just mogged this tame impala test.
🌀 I gave 4 fast models the same Currents cover. Same prompt. One shot via aimlapi.
🌀 I gave 4 fast models the same Currents cover. Same prompt. One shot via aimlapi.
My own agent/harness running entirely on my own hardware. It's the Qwen 3.8 Flash model running on one Spark, plus all the voice models running on a…
Cruz shipped a custom ExLlamaV3 fork and a native engine path. I ran the same model on the same Spark: 102.6 max, 80 tok/s on code, 70 tok/s on…
Since I have my own quantized version of that model thatI'm testing, ~12gb file size, I decided to throw Bonsai2 in there also to get a
That means AI can finally understand what was said AND what happened on screen at the same time. Here’s how I’d test it: → Open the Qwen chat app.
Anyway related, here is harness on a rented NVIDIA H200 141GB optimizing training of QWEN-3.8-27B on a complex reasoning task -runs a few short…
Build and automate anything with the 2.4T parameter open weight model.
TokenHarbor also gave me $12 in free credits - Qwen 3.8 Flash - DeepSeek V4.1 Flash - $12 free credits claimed Use free here: base URL: use it while…
👀 Qwen 3.8 Omni Flash can WATCH the video, HEAR the audio, find the important moments, skip the fluff, and keep up with insanely long context.
Want the SOP? DM me. 💬
Alibaba just dropped an AI that watches video like a human. Sees the screen. Hears every word. At the same time. It's called Qwen 3.8 Omni Flash.
I gave the same Voxel Pagoda Garden prompt to all 3 (The EXL3 4.05 bpw variant on my dgx spark) and the outcome was interesting, yet re-affirming
🚀 Gemini 3.8 Live Qwen 3.8 Omni Flash Qwen 3.8 Live Translate Bonsai 2 27B Needle 3 Jev Laya Nimble Occamy ZGCM Meridian R2T2 Jing Dao Dream RSI…
🔹 Qwen 3.8 Omni Flash: 9/10 · $0.013 · smooth gameplay and cheaper 🔹 DS V4.1 Flash: 9/10 · $0.0089 one-shot with the lowest cost 🔹 Gemini 3.8…
$2/hr if you need some quick work done
This is a full breakdown of Splash, the new open source local inference engine from inco_ai , built specifically around two models: Qwen 3.8 27B and…
Starting to get it dialed in a bit better, it can handle very complex instructions. Having the ability to run this locally is freaking amazing!
Check out the 2x2 panel of a level 1 Priest! Built in Three.js with Hermes Agent.
Qwen 3.8 27B on cerebras reads the battlefield, OpenCV reads cards + elixir, and Jev picks the card + placement. fresh account, starter deck.