[Announcement] On hiatus until August 21

The open models are landing thick and fast right now, and my gallbladder picked exactly this moment to give out. Nothing new here for a couple of weeks.
Hi everyone. This is a genuinely great stretch to be into local models — the full DeepSeek V4 Flash, Qwen 3.8, MiniMax H3, all within days of each other. There is a ridiculous amount to play with. Go enjoy it.
I'd love to be joining in. Instead I've spent about a week in a hospital bed: gallstones, a blocked duct, and the kind of inflammation that gets you admitted rather than sent home with painkillers. I kept telling myself I could still get some benchmarking done between the bad stretches.

I can't even.
(Anyone who can actually run benchmarks through this is built differently. XD)
Look after yourselves out there.
Read next
- 2026-08-07[Benchmark] Running MiniMax-H3 on a DGX Spark — and why NVIDIA VSR is off the table for now
Fifteen seconds of 1080p video with audio in 741s on a GB10 DGX Spark. Swapping Real-ESRGAN for SPAN saved 282s, and nvidia-vfx ships x86_64 wheels only — nothing for ARM.
- 2026-08-06[Just for Fun — Advanced] Wait, a 2080 Ti Can Run MiniMax-H3? 1080p With Audio on a 2018 Card
The four files are 38 GiB on disk; the card has 22. A modded 2080 Ti 22G still renders 15s of 1080p with audio in 23 minutes. Full config, measured speed and quality, then how it got there.
- 2026-08-06[Benchmark] Twice as Fast: MiniMax-H3 on an RTX 5090, 625s Down to 314s
Three stacked changes took a 15-second 1080p MiniMax-H3 render from 625s to 314s on one RTX 5090: 14 steps, SageAttention 2.2.0, and RTX VSR replacing Real-ESRGAN.
- 2026-08-04[Benchmark] Running MiniMax-H3 on one RTX 5090: four files, 31.7 GB, 175s per talking clip
Beginner walkthrough for MiniMax-H3, the 33B model that generates video and stereo audio in one pass. Full precision is 115 GB; quantized it fits one RTX 5090. What to download, where it goes, how to prompt it.
Don't miss the next one
Subscribe, and you won't.
One-click unsubscribe anytime.