AI Tools.

Search

image text to video models

1 models · ranked by HuggingFace downloads

MiniMax-H3

by MiniMaxAI

MiniMax-H3 generates synchronized audio-video clips from text or image prompts, producing output with coherent ambient sound and motion together. It supports multiple input modalities including text-to-video, image-to-video, and video-to-video transformation pipelines. The model ships on Diffusers and uses safetensors checkpoints, making it straightforward to integrate into ComfyUI or custom generation workflows.

5,092,067 ↓ · 4,876 ♡