FRAMEWIREIndonesiaUpdated Sep 9Live wire
0:00 / 0:00

Your 200 tokens per second is the new batch mode.

Cerebras CTO Sean Lie, the day after Hot Chips: what used to count as fast, 100 or 200 tokens per second, is quickly becoming batch. Fine for prompt processing. Fine for parallel jobs. Not fine for an agent loop where you wait on

godgivenSep 91
0:00 / 0:00

DeepSeek-V4.1-Flash made this

Fede(URU) 🇺🇾Sep 9
0:00 / 0:00

DeepSeek-V4.1-Flash did it again

One of the most prettiest Three.js spiral galaxy

Fede(URU) 🇺🇾Sep 9
0:00 / 0:00

Stop downloading LLMS your machine was never going to run

Llmfit is an open-source tool that scans your hardware first, then tells you exactly which models will actually run well on your setup. It inspects: → RAM → CPU → GPU(s) → VRAM / unified memory Then it scores every

Charly Wargnier ♨️Sep 912