MiniMaxAI/MiniMax-H3
Image-Text-to-Video • 33B • Updated • 1.43k
Video generation with a synchronized soundtrack
Note Text-to-video, Image-to-video
Unquantized MiniMax-H3 from image, audio, video refs
Note Reference to video (image-reference, audio-reference - including lipsync, video reference)
Qwen3-VL 33B prompt conditioning as a service