FreeVideo runs MiniMax H3 on 8 GB GPUs

4 10 2026
FreeVideo runs MiniMax H3 on 8 GB GPUs

FreeVideo is a local inference engine for the MiniMax H3 video model that operates inside ComfyUI or via a Linux command line. By combining the 8-step VDN-H3 model with weight streaming, asynchronous prefetching, and hardware-adaptive FP8 execution, it runs on consumer hardware with as little as 8 GB of VRAM and 16 GB of system RAM. It accepts text, image, video, and audio inputs while supporting community LoRAs.

It is worth testing if you want to experiment with local video generation without renting cloud instances, provided you have patience for streaming latency on lower-end hardware. The setup handles the memory management well, but generation speeds on an 8 GB card will still be slow enough to rule out rapid iteration. Skip it if you are looking for real-time preview workflows.

more: https://github.com/FlashML-org/FreeVideo


Actions

Information

Leave a comment