FRAMEWIREIndonesiaUpdated Aug 30Live wire
0:00 / 0:00

Qwen 3.8 takes a little longer to think things through...

Also Qwen 3.8:

Yago JesusAug 291
0:00 / 0:00

This is an example of a work created by a local Creative Agent on Qwen-3.8-Next-Flash, starting from a single picture.

The referenced picture is included in the first frame, and it has a loop configuration that returns to it at the end. I purposely do not have any say in the production, including the development.

Nobu-Kobayashi : Generative AI TechnologyAug 2986
0:00 / 0:00

DeepSeek V4 Flash • Qwen 3.8 Flash

GLM 5.3 Flash • Gemini 3.7 Flash

Fabiano FirmoAug 292
0:00 / 0:00

262K context. On a 16GB RTX 5070 Ti.

🤯 Qwen 3.8 27B Q3 hits ~25 tok/s while an adaptive llama.cpp fork streams KV cache between RAM ↔ VRAM. Stock llama.cpp starts thrashing around ~120K context. Same consumer GPU. 2x+ the usable context. This could be huge for local LLMs.

SachinAug 293