Shoutout to 0xBakeer for launching his own inference engine TandemLLM !
! ! I've tested it in the last few hours and I can say Qwen 3.8 27B on a DGX Spark is much more snappier when run with it. Also the included inference monitoring dashboard is very informative and
AITuber OnAir Core has been updated. Models released in the past few days, such as GPT-6.1 Sol and Eleven v4, can now be used together.
・LLM: GPT-6.1 Sol, GLM-5.3, Claude Opus 5.5, Sonnet 5.5, Grok 4.7, Qwen 3.8 of the OpenRouter operating system ・TTS: Eleven v4, Cartesian Sonic 3.6, Deepgram
Qwen 3.8 Flash Next made a voxel Middle‑earth zoom chain + a 15 s orbit clip.
Every frame came off one RTX 3090. How the agent does it without me touching a keyboard. The interesting constraint: my agent's brain (Qwen3.8‑Flash‑Next IQ3, llama.cpp, ~21 GB) is resident on the
How to get started with open AI models (with
I have recorded a simple video for all of you who ask us. 1. How to enter NaN 2. What NaN is like inside 3. How to configure your api-key Pum, pum, pum! Quick, simple and for the whole family.
You have reached the end of the archive
All of qwen38