Used Qwen 3.8 27b for it all (local) with OMP.
I love making silly things like this, and AI allows me to move so much quicker (and with better art). Link below.
こちらがQwen-3.8-Next-Flash上のローカルCreative Agentに1枚絵から自由に作らせた作例。
参照させた絵は1フレーム目にはいっていて、最後にそこに戻るループ構成になっている。展開も含めた演出にはあえて口を出していない。
DeepSeek V4 Flash • Qwen 3.8 Flash
GLM 5.3 Flash • Gemini 3.7 Flash
262K context. On a 16GB RTX 5070 Ti.
🤯 Qwen 3.8 27B Q3 hits ~25 tok/s while an adaptive llama.cpp fork streams KV cache between RAM ↔ VRAM. Stock llama.cpp starts thrashing around ~120K context. Same consumer GPU. 2x+ the usable context. This could be huge for local LLMs.
You have reached the end of the archive
All of qwen38