Generating AI Video Locally with MiniMax H3 in ComfyUI on an RTX 3090
Melvin Vivas · X video post · 2026-08-07 · 0:15 · 290 views · Open on X
Topics: AI Dev Tools & Productivity, Industry Trends & Job Market · Level: intermediate
Summary
Melvin Vivas shares a 15-second war-scene clip he made on his own computer with the MiniMax H3 video model, running in ComfyUI on one RTX 3090 GPU. He says the main appeal is that it's free to run locally. He also gives real numbers: 22 minutes to generate the clip at a low test resolution of 864x480. He likes how closely the model follows prompts. The post is a quick hands-on report, not a tutorial.
Key points
- The MiniMax H3 video model can run locally in ComfyUI on a consumer RTX 3090 GPU (24 GB of VRAM).
- Running it locally is free, which the creator gives as his main reason for liking it.
- A ~15-second clip took about 22 minutes to generate on the RTX 3090.
- He kept the resolution low (864x480) while testing prompts, then plans to go higher once a prompt works.
- He says the model follows prompts well (good prompt adherence).
- The sample clip is an action/war scene with spoken dialogue ("Move! Off the boat!"), which suggests the output can include audio.
Resources mentioned
- MiniMax H3 · tool · huggingface.co · free
MiniMax video generation model that turns a starting image and an audio track into a lip-synced video.
Also in: MiniMax H3 Image + Audio-to-Video Lip-Sync Demo (with Irodori-TTS v3) (Melvin Vivas on X · notes), Running MiniMax H3 Locally on a Single RTX 3090 (Demo) (Melvin Vivas on X · notes) - ComfyUI · tool · x.com · free
The official X account for ComfyUI, the node-based tool for generative AI workflows, which announces things like ComfyUI Agent.
Also in: ComfyUI Agent: early access announcement (Melvin Vivas on X · notes), Comfy Router: One API for Image, Video, 3D and Audio Model Providers (Melvin Vivas on X · notes), Running MiniMax M3 Locally in ComfyUI on an RTX 3090 (Melvin Vivas on X · notes), Demo: Real-Time VFX Compositing Inside ComfyUI (Melvin Vivas on X · notes) and 23 more - NVIDIA GeForce RTX 3090 · tool · nvidia.com · paid
Consumer GPU with 24 GB of VRAM, used here to run a long-context LLM locally.
Also in: Portable Computer Now Runs AI Agents and Models Locally on NVIDIA RTX PCs (Melvin Vivas on X · notes), Long-Context Local LLM on an RTX 3090: .env Config and ~64 tok/s Benchmark (Melvin Vivas on X · notes), Running Ornith-1.5-35B-A3B (Q4_K_M) at 128k Context on an RTX 3090 with llama.cpp (Melvin Vivas on X · notes), Running MiniMax H3 Locally on a Single RTX 3090 (Demo) (Melvin Vivas on X · notes) and 2 more
Try this
- Test prompts at low resolution (e.g., 864x480) first to save time, then render at higher resolution.
- Set up a local text-to-video workflow in ComfyUI with MiniMax H3 on a 24 GB GPU and measure how long generation takes at different resolutions.
More in AI Dev Tools & Productivity
- AI One-Shot Rust Rewrite of TerminalTextEffects: 43x Faster Startup
- Use Codex to set up and fix local tooling
- MiniMax H3 Image + Audio-to-Video Lip-Sync Demo (with Irodori-TTS v3)
- Running MiniMax M3 Locally in ComfyUI on an RTX 3090
- Cursor Agents Get Google Workspace Plugins
- Record a skill in Claude Desktop