On my Strix Halo PC, Qwen 3.8 Flash-Next on Halogen passed 66/72 coding jobs and generated at 38.0 tok/s versus 13.9 in my earlier setup.
Different quant, backend and drafting make this a configuration comparison, not an engine-only test. Runs:
Different quant, backend and drafting make this a configuration comparison, not an engine-only test. Runs:
The model is soo tasteful you can feel it in responses. here are the data and specs: qwen 3.8 flash next, 180b moe, official fp8 45 tok/s fresh, 35…
But as context grew, generation fell from about 15 to 8 tok/s across three harnesses.
The speed is absolutely CRAZY 🤯 Local AI is no longer a compromise.
Ik teste het met Qwen 3.8 27b op mijn laptop! De opdracht: maak een robot met uitwisselbare onderdelen, 4 animaties en exploded view.
Made with Tesseract by Mirage Model - Qwen 3.8 27b Hardware - Single Nvidia 96 RTX 6000 Workstation Edn Rate my edits... today...yikes
You can access it for free and as an open-source project at qoder_ai_ide
1M free tokens to play with models like: - Claude Opus 5.5 - Claude Fable 5.1 - GPT-5.6 Sol - GPT-5.6 Luna - Grok 4.6 - Qwen 3.8 Max - DeepSeek V4…
Het model draaide lokaal, bouwde zelf de website en genereerde de afbeeldingen via api. Ongeveer een half uurtje werk, met resultaat!
Pasar Model Vikey Top 5 model paling laris di Vikey: 🥇 Qwen 3.8 27B Uncensored: jawabnya lugas, no ceramah 🥈 DeepSeek V4 Flash: input cuma Rp1.080…
⊘⬡ 🜍 🜂 A Leaderboard Is Not the Country Full Gate-Empty lens on where China’s AI sits · holes named, not averaged TN-CHINA-AI-WHERE-π-001 rev.b…
Geef Qwen 3.8 27b een lijst met je producten, barcodes en prijzen, en binnen een half uur kun je scannen en facturen sturen!
EXL3 - Qwen 3.8 Flash next - single spark. <100k tokens decode total.
Ok here's my voxel pagoda "Thai Island Paradise" - Qwen 3.8 Flash Next using the EXL3 engine (credit ViC305 ) which I patched for Hermes…
I love botting. i’m going through and improving the gym and world model. qwen 3.8 27b on my rtx5090 with jev for fighting.
Current OpenClaw doing a 7sec run with Sol as primary, Luna as light and Jev (or Qwen 3.8 Flash Next just 2sec slower) searching for a flight in a…
How did Opus 5.5, GPT 6 Sol and Grok 4.7 etc. stack up? Let's find out!
GLM 5.3 Prime is the high-speed variant of GLM-5.3, live today on OpenRouter.
No catch found yet. Qwen 3.8-Flash. 1M token context, enough to feed it an entire codebase or a full-length book in one shot.
Qwen 3.8 Flash plans the build, then hands it to Qwen 3.8 27B, all running locally across the GPU, CPU and 128GB of unified memory on one AMD Ryzen…