Free GPUs 2x NVIDIA T4s & run Qwen 3.8 27B at 14 t/s with 120k context
120k context on FREE Kaggle T4s. - FP16 KV - 14 t/s - No credit card. - No expensive GPU. - Total VRAM used: only 26.6 GB across both cards. - Under 5 minutes setup. - Kaggle gives you 30 hours/week of
Preliminary experiment with Qwen 3.8/MiniMax H3
I'm developing a prompt that takes a comic page, cuts it out and turns it into a video.
Running Qwen 3.8 27B on 2×5090s with a few basic optimizations
~230 tokens/s (video is real-time) two years ago building a pubsub would have taken me an entire day now it takes 8 seconds
WARNING: Next week could be crazy for AI.
A list that is circulating points to 11 possible new models and updates: 🤖 Claude Fable 5.1 🚀 GPT-6 Astra ⚡ Grok 4.7 🇨🇳 GLM-5.3 Flash 🔓 Qwen 3.9 Open Weights 🎬 Composer 3 🧠Ox Alpha Flash ✨ Gemini 3.8
You have reached the end of the archive
All of qwen38