I have run a test on Token/Second specifically running Deepseek V4 Flash in Deepseek Harness
Runinfrai claims 250 token/sec while CommandCodeAI have no claims about speed, yet I saw really similar results. RunInfra: 2 turns · 73 steps · TTFT avg 2.8s · 108 tok/s cache hit 90%
Top AI LLMs for Every Task
Try Gemini 3.1 Pro for FREE 👉 → Writing & Research: GPT-5.4, Claude 4.6, Gemini 3.1 Pro, Perplexity → Social Content: Grok 4, GPT o3, DeepSeek → Academic / STEM: Claude Opus 4.6, MiniMax M2.7 One place to access the AI
My minimax h3 skills are evolving.
This is low quality config run with Deepseek-v4-flash co-tenant + Hermes orchestrating ComfyUI through their OSS MCP server. Took 33-minutes to create 5 x 5s clips with persistent voice and character. Hermes sticking the videos together.
Made another long-form video for my youtube channel.
I am getting the hang of it a bit. But it is still a lot of work to put together a decent video. And I am no where near to the quality of videos that I watch by other creators. I am still getting there. Feel free support my
You have reached the end of the archive
All of deepseek