MiniMax-H3 arrives on HuggingFace with multimodal generation capabilities including video at up to 2K resolution and 15-second duration
Read the original at old.reddit.com→MiniMax H3 is a general-purpose, omni-modal generative system. It supports unified understanding of multimodal contexts composed of text, images, video, and audio, and can generate video with native stereo audio at...
Original headline: "MiniMax-H3 now on huggingface"
Coverage timeline
- Aug 3, 03:06 UTC r/LocalLLaMA lead source MiniMax-H3 now on huggingface