Gate.AIBlogHailuo 2.3: Complete Specifications, Pricing, API Access & Use Cases (2026)

    Hailuo 2.3: Complete Specifications, Pricing, API Access & Use Cases (2026)

    Models

    What Is Hailuo 2.3?

    Hailuo 2.3 is MiniMax’s short-form video generation model, released on October 28, 2025, featuring text-to-video and image-to-video generation, 6–10 second output options, and 768p or 1080p resolution choices, with Gate.AI listing pricing by duration and resolution as of July 2026. MiniMax introduced Hailuo 2.3 as an update to Hailuo 02 and described improvements in dynamic expression, physical actions, stylization, character micro-expressions, and response to motion commands.

    Hailuo 2.3 is not a large language model for chat, reasoning, or coding. It is a generative video model for producing short video clips from text prompts or source images. This makes it relevant for creative teams, social media producers, e-commerce marketers, storyboard artists, and developers building short-form video generation workflows.

    MiniMax’s official video API documentation describes video generation as an asynchronous process: create a task, poll the task status, and retrieve the generated file after completion. Gate.AI’s public documentation also describes an asynchronous video generation endpoint with task submission, polling, billing fields, and download URL retrieval.

    What Are Hailuo 2.3’s Key Specifications and Pricing?

    The following table separates provider-verified information from Gate.AI listing details. MiniMax’s official pricing documentation uses video points, while the Gate.AI listing for this model presents dollar pricing by resolution and duration. MiniMax states that MiniMax-Hailuo-2.3 deducts 1 video point for a 768p 6-second video, 2 video points for a 768p 10-second video, and 2 video points for a 1080p 6-second video.

    Field Verified Value
    Provider MiniMax, as of July 2026
    Model Family Hailuo video generation models, as of July 2026
    Model Type Short-form AI video generation model, as of July 2026
    Release Date October 28, 2025, as of July 2026
    Context Window Not applicable; Hailuo 2.3 is a video generation model, not a token-context LLM, as of July 2026
    Prompt Limit Gate.AIdocumentation states Hailuo prompt limits vary by upstream provider and lists Hailuo at 2000 characters, as of July 2026
    Input Pricing No token-based input price confirmed; video generation is billed by task specification, duration, or resolution, as of July 2026
    Output Pricing As perGate.AIlisting: 768p/6s \$0.28, 768p/10s \$0.56, 1080p/6s \$0.49; 1080p/10s not listed, as of July 2026
    Pricing Unit Per generated video, based on duration and resolution, as of July 2026
    MiniMax Official Pricing Unit Video points, with deductions varying by model, duration, and resolution, as of July 2026
    Modality Support Text-to-video and image-to-video, as of July 2026
    Supported Input Types Text prompts and first-frame image input for image-to-video workflows, as of July 2026
    Supported Output Types Generated video file, as of July 2026
    API Access MiniMax API andGate.AIasynchronous video API, as of July 2026
    Model ID MiniMax official model ID: MiniMax-Hailuo-2.3;Gate.AIlisting model ID: minimax/hailuo-2.3, as of July 2026
    Availability MiniMax states Hailuo 2.3 was rolled out across Hailuo AI, mobile app, and Open Platform API, as of July 2026
    Knowledge Cutoff Not applicable to this video generation model; official training data cutoff not specified, as of July 2026
    Rate Limits MiniMax video packages list RPM tiers by package; exact account limits may vary, as of July 2026
    Fine-tuning Support Not confirmed from official sources as of July 2026
    Streaming Support Not applicable to the documented asynchronous video task workflow, as of July 2026
    Batch API Support Not confirmed from official sources as of July 2026
    Tool / Function Calling Not applicable to this video generation model, as of July 2026
    Structured Output / JSON Mode Not a generation feature; API task responses use structured JSON, as of July 2026
    License / Usage Restrictions Not fully specified in the available model documentation as of July 2026

    What Can Hailuo 2.3 Do That Makes It Useful in Production?

    Hailuo 2.3 can generate short videos from text prompts, making it useful for fast visual ideation, campaign mockups, storyboard tests, and short-form creative drafts. MiniMax’s video generation workflow supports asynchronous generation, which fits production systems that submit tasks, poll status, and retrieve completed media files.

    The model also supports image-to-video generation. MiniMax’s image-to-video API documentation shows a workflow where a prompt, first-frame image, model ID, duration, and resolution are used to create a video task. This is useful when a team already has an approved product image, character frame, concept still, or brand visual and wants to add controlled motion.

    MiniMax describes Hailuo 2.3 as improving dynamic body movement, motion command following, facial micro-expression rendering, stylization, lighting behavior, and object motion response. These capabilities matter in short video production because motion consistency, subject stability, and camera-control adherence often determine whether generated clips are usable without excessive reruns.

    Hailuo 2.3 may be especially relevant for AI short-video generation workflows where teams need compact clips, predictable pricing, and prompt-controlled camera movement. Related model profiles such as Seedance 2.0 video generation can help to compare different video model families.

    What Are Hailuo 2.3’s Supported Modalities?

    Modality Supported? Notes
    Text Input Yes Used for text-to-video prompts describing subject, scene, motion, camera, and style.
    Image Input Yes Used as first-frame input in image-to-video workflows.
    Video Input Not confirmed for Hailuo 2.3 standard Gate.AI’s generic video API supports reference media fields, but Hailuo 2.3 standard video-input support is not confirmed in the model facts.
    Audio Input Not confirmed for Hailuo 2.3 standard Generic video APIs may include audio-related fields, but audio input is not confirmed as a Hailuo 2.3 standard capability.
    Video Output Yes Output is a generated video file retrieved after task completion.
    Audio Output Not confirmed for Hailuo 2.3 standard Gate.AI’s generic video task schema includes generate_audio, but Hailuo 2.3 audio output is not confirmed in the available model listing.

    Where Does Hailuo 2.3 Fall Short?

    Hailuo 2.3 is designed for short-form video, not long-form scene continuity. The Gate.AI listing covers 6-second and 10-second generation options, while MiniMax’s official pricing documentation lists point deductions for 768p/6s, 768p/10s, and 1080p/6s. For 1080p/10s, the Gate.AI listing does not provide a price, so teams should verify availability before building fixed-cost workflows around that configuration.

    Pricing also requires careful source separation. MiniMax official documentation uses video points, while the Gate.AI listing uses dollar prices for specific duration-resolution combinations. These are separate billing presentations and should not be merged unless a pricing page explicitly maps one unit to the other.

    This is a general AI limitation and is not model-specific unless stated by the provider: generated videos may contain artifacts, visual distortions, unrealistic physics, identity drift, prompt misunderstanding, or unexpected object changes. Human review is necessary before using generated content in brand, legal, political, medical, financial, safety-sensitive, or rights-sensitive environments.

    Copyright and likeness review are also important for AI video generation. Production teams should avoid prompts that request protected characters, real-person likenesses without permission, logos, copyrighted scenes, or other restricted assets unless they have the necessary rights and policy clearance.

    What Is Hailuo 2.3 Best Used For?

    Use Case Why Hailuo 2.3 May Fit Important Limitation
    Short social video concepts 6–10 second outputs fit short-form creative testing and visual iteration. Not designed for long narrative continuity.
    Text-to-video storyboards Prompt-based generation helps teams test scenes, motion, and composition quickly. Prompt adherence can vary by scene complexity.
    Image-to-video animation First-frame image input helps anchor product, character, or brand visuals. Subject consistency still requires review.
    Camera movement tests Gate.AIprompt guidance includes subject, motion, scene, camera, and style fields. Fine camera control may need multiple attempts.
    E-commerce creative drafts MiniMax highlights object motion response and e-commerce ad examples in its release note. Final ad use needs rights, accuracy, and brand review.
    Creative generation pipelines Asynchronous API design fits systems that queue, poll, and retrieve video tasks. Latency and success rates should be tested per workflow.

    How Does Hailuo 2.3 Compare to Seedance 2.0 and Veo 3.1?

    Comparison Area Hailuo 2.3 Seedance 2.0 Veo 3.1 Scenario Fit
    Provider MiniMax ByteDance Seed Google DeepMind / Google Choose based on access channel, rights policy, and workflow needs.
    Core Use Short text-to-video and image-to-video generation Multimodal audio-video generation and editing Video generation with native audio and cinematic controls Hailuo 2.3 fits compact video tasks; Seedance 2.0 and Veo 3.1 fit broader multimodal workflows.
    Inputs Text and image confirmed Text, image, audio, and video Text, image, and reference-guided workflows Hailuo 2.3 is simpler; competitors may offer richer reference control.
    Output Duration 6–10 seconds inGate.AIlisting Up to 15 seconds in official launch material 8 seconds in Gemini API documentation Duration needs should guide selection.
    Resolution 768p and 1080p options inGate.AIlisting Official launch emphasizes high-quality audio-video output; exact API resolution should be checked per channel Gemini API documentation lists 720p, 1080p, or 4K for Veo 3.1 Confirm resolution through the exact platform used.
    Audio Not confirmed for Hailuo 2.3 standard Audio-video joint generation emphasized Native audio emphasized Use Seedance 2.0 or Veo 3.1 when synchronized audio is a central requirement.
    Production Fit Short-form visual ideation and image-to-video animation Multimodal creative production Cinematic clips with native audio and reference controls No universal winner; choose by modality, duration, price, and access constraints.

    Seedance 2.0 and Veo 3.1 are useful comparison models because they target modern AI video generation with stronger multimodal and audio-video workflows. ByteDance states that Seedance 2.0 supports text, image, audio, and video input modalities and up to 15-second high-quality multi-shot audio-video output. Google describes Veo 3.1 as supporting text-to-video, image-to-video, text-to-audio-plus-video generation, and realistic physics, with Gemini API access for 8-second videos with native audio.

    How Do I Access Hailuo 2.3 Through Gate.AI?

    As per Gate.AI listing, Hailuo 2.3 is available with the model ID minimax/hailuo-2.3. Gate.AI’s public documentation verifies the asynchronous video task API at POST /api/v1/videos, Bearer-token authentication, optional idempotency keys for safe retries, video prompt fields, duration, resolution, aspect ratio, reference media fields, webhook support, and task polling through GET /api/v1/videos/{job_id}.

    Gate.AI’s pricing page states that video and other multimodal capabilities are billed by generation count, duration, resolution, or task specification. For Hailuo 2.3 specifically, the Gate.AI listing shows 768p/6s at \$0.28, 768p/10s at \$0.56, and 1080p/6s at \$0.49 as of July 2026.

    Python Example

    1. import os
    2. import uuid
    3. import requests
    4. api_key = os.environ["GATE_AI_API_KEY"]
    5. headers = {
    6. "Authorization": f"Bearer {api_key}",
    7. "Content-Type": "application/json",
    8. "Idempotency-Key": str(uuid.uuid4()),
    9. }
    10. payload = {
    11. "model": "minimax/hailuo-2.3",
    12. "prompt": (
    13. "A cinematic product shot of a matte black electric scooter "
    14. "turning slowly under neon city lights, smooth camera pan, "
    15. "stable subject, realistic reflections."
    16. ),
    17. "duration": 6,
    18. "resolution": "1080p",
    19. "aspect_ratio": "16:9",
    20. "generate_audio": False
    21. }
    22. response = requests.post(
    23. "https://api.gate.ai/api/v1/videos",
    24. headers=headers,
    25. json=payload,
    26. timeout=60
    27. )
    28. response.raise_for_status()
    29. job = response.json()["data"]
    30. print("Job ID:", job["job_id"])
    31. print("Status URL:", job["status_url"])

    curl Example

    1. curl https://api.gate.ai/api/v1/videos \
    2. -H "Authorization: Bearer $GATE_AI_API_KEY" \
    3. -H "Content-Type: application/json" \
    4. -H "Idempotency-Key: $(uuidgen)" \
    5. -d '{
    6. "model": "minimax/hailuo-2.3",
    7. "prompt": "A cinematic product shot of a matte black electric scooter turning slowly under neon city lights, smooth camera pan, stable subject, realistic reflections.",
    8. "duration": 6,
    9. "resolution": "1080p",
    10. "aspect_ratio": "16:9",
    11. "generate_audio": false
    12. }'

    To check generation progress, use the returned job_id with GET /api/v1/videos/{job_id}. Gate.AI’s documented response includes status, model, status URL, download URL after completion, duration, resolution, estimated cost, billed cost, billing status, and timestamps.

    FAQs

    What is Hailuo 2.3?
    Hailuo 2.3 is MiniMax’s short-form AI video generation model for text-to-video and image-to-video workflows. It is designed for short creative clips, motion-controlled prompts, subject stability, stylized visuals, and image-based video animation.

    How much does Hailuo 2.3 cost on Gate.AI?
    As per Gate.AI listing, Hailuo 2.3 costs \$0.28 for 768p/6s, \$0.56 for 768p/10s, and \$0.49 for 1080p/6s as of July 2026. A 1080p/10s price is not listed.

    Can developers access Hailuo 2.3 through an API?
    Yes. MiniMax documents video generation through its API, and Gate.AI documents an asynchronous video generation endpoint. As per Gate.AI listing, the model ID is minimax/hailuo-2.3.

    When should teams choose Hailuo 2.3 instead of Seedance 2.0 or Veo 3.1?
    Hailuo 2.3 may fit short text-to-video and image-to-video workflows where duration, resolution, and iteration cost are central. Seedance 2.0 or Veo 3.1 may fit projects that require stronger audio-video or multimodal reference workflows.

    The content herein does not constitute any offer, solicitation, or recommendation. You should always seek independent professional advice before making any investment decisions. Please note that Gate may restrict or prohibit the use of all or a portion of the Services from Restricted Locations. For more information, please read the User Agreement

    Related Articles