Basic circuit design with my FLAI harness and using qwen 3.8 flash via open router
Qwen 3.8 27b NVFP4 + 5090 + Podman + Ninfer + arch hitting 443 tok/s decode @ c4
Oh. my. god.
Cosa è in grado di fare il mini PC AMD Strix Halo?
È pensato per far girare i modelli AI in locale. Cioè sul tuo tavolo, in azienda, non sui server di qualcun altro. Lavora sui tuoi dati senza spedirli a nessuno e non si ferma quando scade un abbonamento. Ci ho fatto girare
Everyone's been complaining Qwen 3.8 27B overthinks and is too slow, so we fixed it.
Introducing UkisAI Swift-27B We trained Swift by reducing model anxiety, the results: 58.3% less token usage, 1.95x faster outputs, <1% loss in accuracy It’s the same performance, just
You have reached the end of the archive
All of qwen38