APIXO has not yet published a MiniMax H3 model page or operational API contract. Route availability, authentication path, request fields, and delivery behavior should remain unpublished until APIXO documents them.
MiniMax H3
MiniMax H3 is a general-purpose multimodal video model that combines text, image, video, and audio context in one generation workflow. It creates videos up to 2K and 15 seconds with native stereo sound, while supporting first-and-last-frame control, mixed-media references, motion transfer, multi-shot storytelling, and instruction-guided video editing.
Inputs
Text / Image / Video / Audio
APIXO Price
Announced at launch
Resolution
768P / 2K
Duration
4-15 Seconds
Audio
Native Stereo Sound
Create with MiniMax H3
MiniMax H3 is not open for generation yet
This page is live ahead of launch. Creation controls, parameter tiers, and pricing appear here the moment MiniMax H3 opens on APIXO — no other page to bookmark.
MiniMax H3 Capabilities
Unified Context
Combine written direction with image, video, and audio references in one request. Assign each source a role such as appearance, motion, camera, voice, rhythm, environment, or style.
First & Last Frames
Define an opening frame, an ending frame, or both to guide transformations, transitions, product reveals, and storyboard beats while preserving a clearer visual trajectory.
Mixed-Media References
Use reference images, videos, and audio together to communicate character identity, product cues, movement, camera language, vocal performance, music, or ambience.
Native Stereo Sound
Generate voice, sound effects, ambience, and music together with the video as native stereo audio rather than assembling isolated audio tracks after generation.
Native Multi-Shot
Build short sequences with coordinated shot changes, scene progression, action continuity, and audiovisual pacing from one prompt-led or reference-guided workflow.
Editing & Motion Transfer
Transform existing visual material through natural-language direction or transfer motion from video references while reviewing identity, composition, occlusion, and source-detail fidelity.
Production concepts suited to MiniMax H3
Audiovisual Product Stories
Combine product imagery, written direction, motion references, and sound intent to develop campaign concepts. Review labels, logos, materials, dimensions, and product geometry before approving generated commercial assets.
Reference-Led Performances
Establish a recurring subject with visual references, then guide movement, camera behavior, voice, or scene context through additional media. Evaluate identity continuity, anatomy, expression, and performance timing across the complete result.
Narrative Previsualization
Translate a treatment into short scenes that combine action, camera language, ambience, and audiovisual pacing. Use generated footage to evaluate ideas before committing to physical production or detailed post-production.
Footage Transformation
Explore style, environment, subject, or scene changes through natural-language instructions and optional references. Inspect temporal boundaries, occlusion, unintended modifications, and source-detail preservation frame by frame.
MiniMax H3 Model Specifications
Verified upstream capabilities to confirm against the final APIXO deployment.
Text / Image / Video / Audio
Input
T2V / I2V / Reference
Workflows
768P / 2K
Resolution
4–15 Seconds
Duration
Native Stereo
Audio
6 Ratios + Adaptive
Framing
MiniMax H3 availability and integration guidance
Notes
APIXO has not publicly documented an operational MiniMax H3 route, model ID, mode enum, or delivery contract.
Do not assume that “MiniMax H3” and “Hailuo 3.0” are interchangeable APIXO names until the route mapping is published.
APIXO duration, resolution, aspect-ratio, reference-count, prompt, and media-file limits remain unpublished.
Do not transfer parameter values, quality tiers, output rules, or billing behavior from Hailuo 2.3 or Hailuo 2.3 Fast.
References guide generation but do not guarantee exact identity, typography, logos, materials, or product geometry.
Review anatomy, hands, object contact, rapid motion, occlusion, camera transitions, speech, and audiovisual synchronization before release.
Frequently Asked Questions
Public launch material describes prompt-led video, image-guided generation, multimodal reference use, audiovisual creation, and natural-language video revision. APIXO may expose only a subset of these workflows.
No APIXO-specific resolution or duration values are currently verified. Values shown by upstream demonstrations, community workflows, or third-party hosts must not be presented as APIXO guarantees.
MiniMax H3 is announced as an audiovisual model with audio-aware generation. APIXO-specific audio inputs, output behavior, channel configuration, sound switches, and file constraints remain unpublished.
No. Mixed-media references communicate creative intent but do not guarantee exact identity, motion, voice, product structure, or frame-level reproduction. Production use requires visual and audio review.
Explore Other Models
Discover more AI models for your next creative workflow