MiniMaxMiniMax·🎬 Video Generation·New

MiniMax H3

anonymized
Try on Venice.ai ↗
Quick reference
MiniMax H3 — TLDR
  • 🆕 MiniMax's newest-generation video model line, released July 2026.
  • 👁️ Image-to-video variant: animates a supplied still into motion.
  • 🏢 Described by MiniMax as an open general-purpose multimodal video model.
  • 🎯 Shares generation with sibling text-to-video and reference-to-video endpoints.
  • 🔧 Routed as a separate catalog entry per input modality.
  • 📚 Part of MiniMax's wider stack spanning text, speech, music, video.
  • 🌐 Documented on MiniMax's own platform API docs.
💰 Pricing
$0.810 – $2.44
per generation
📅 On Venice since
Jul 30, 2026
2 days ago
Provider

MiniMax is an AI company building generative models across multiple modalities, with a focus that spans both language understanding and audio creation. Their rapid release cadence in early 2026—delivering several new models within just a few months—reflects…

Read full profile →
10 models on Venice
3 text · 3 video · 3 music · 1 tts
Since Feb 12, 2026

About this model

MiniMax H3 (image-to-video) is the picture-conditioned entry point to MiniMax's H3 video generation line, released in July 2026. You supply a still image, optionally with a prompt, and the model produces a moving clip from it. MiniMax's own API documentation announces H3 as a new-generation open general-purpose multimodal video model, rather than a narrowly task-specific generator.

It ships alongside two same-generation siblings that share the H3 name and release date but differ in what they consume: MiniMax H3 builds a clip from a written prompt alone, while MiniMax H3 R2V conditions on reference material to carry a subject, character, or style across into the output. In this catalog the three are listed as separate families purely because their input routing differs.

Compared with MiniMax's earlier video generations, which were published as individual, purpose-built video endpoints, H3 is positioned by the provider as one general-purpose multimodal video system covering the generation surface as a whole. Detailed clip-length, resolution and benchmark figures are not covered here, as no primary MiniMax specification or independent evaluation of H3's video quality is cited.

H3 sits within a broad MiniMax model family: reasoning and coding models such as MiniMax M3 Preview and MiniMax M2.7, speech synthesis in MiniMax Speech-02 HD, and music generation in MiniMax Music 2.6.

This About section is AI-generated from public sources (Claude Opus 5), with no human editing. It may contain inaccuracies — verify critical details against the sources listed above.

Data sources: Venice API · HuggingFace · Wikipedia — enrichment updated 23h ago