FRAMEWIREIndonesiaUpdated Sep 21Live wire
0:00 / 0:00

Qwen 3.8 27B doing an agentic task with subagent at ~120 tok/s on M5 Max MacBook Pro in lmstudio Bionic using inco_ai Splash engine

At launch a little more than a month ago the model was running around ~20tk/s It’s 6x faster now, local AI is progressing at lightning speed!

Adrien GrondinSep 2193
0:00 / 0:00

Like any distilled model, Ternary Bonsai 2 27B inherits refusals and censorship from its teacher model: Qwen 3.8 27B.

Fortunately, there’s a very simple and elegant way to reduce or perhaps even completely eliminate them. 🧵

Private LLMSep 2123
0:00 / 0:00

Your US stock position still carries risk even after your broker closes.

So I built Tesrune for the Bitget_AI S2 Hackathon. Tesrune is a dark-hours AI trading desk for US equities. You add the stocks you're holding. While your broker is closed, Tesrune watches relevant news,

MystiqueMideSep 2161
0:00 / 0:00

Talking about Local Benchmaxxing - here is my 2 year old Intel i9 64 GB machine with Nvidia RTX 4090 running Qwen 3.8 Flash next at 30 t/s.

Happily using opencode at my Macbook Pro using my Older Intel Machine as an API endpoint. Quant : AD-3.84bpw-IQ4_XS-M64

Akash GoswamiSep 21