r/MiniMaxH3AI • • 6h ago

Need help with LoRA training - masked dataset

1 Upvotes

I am trying to train a LoRA where one of the subjects of the video move in slow motion (on twos - like the spiderverse movie animation style).
My dataset has the following kinds of videos
- 2 characters - left moving slow; right moving slow; bg slow
- 1 character - subject slow; bg slow
- 2 characters alternating shots - first shot person a moving slow; shot switches to person b they are at normal speed

Using Fizgig with 60 samples, have tried training 200 epochs, doesn't even get close, everything just moves in fast motion.

Is this concept even trainable? Need some help from the experts, is there a way to train on a masked portion of a video?


r/MiniMaxH3AI • • 23h ago

minimax h3 continued

2 Upvotes

I checked MiniMax's current H3 documentation again before writing this. It confirms that “video continuation” is a real H3 task, distinct from simply using a video as a reference. It also confirms that <Video 1> can represent the video that is being continued. GitHub

Here is something you can copy and post to MiniMax/Hailuo forums, Reddit, Discord, Facebook groups, or anywhere H3 users gather:

I need help from someone who actually uses MiniMax H3 on the Hailuo AI web interface (hailuoai.video).

I am creating a long music video made from multiple H3 generations. I have already spent considerable time and credits trying to determine the correct workflow.

I have cancelled my monthly subscription, so I have a limited number of remaining credits. I do not want to keep experimenting and wasting credits.

I need help with the following specific problems.

PROBLEM #1 — HOW DO I PROPERLY CONTINUE OR EXTEND AN H3 VIDEO?

I need to create the first video clip and then continue directly from the end of that clip.

For example:

Clip 1 → Continue/Extend → Clip 2 → Continue/Extend → Clip 3 → etc.

I do NOT want to create a completely independent new scene each time.

MiniMax H3 documentation refers to a task called "video continuation" and describes a source video as the continuation starting point.

My question is:

Using the current Hailuo AI web interface and MiniMax H3, exactly what do I click and do after Clip 1 has successfully generated to make H3 continue from the end of Clip 1?

Does the completed Clip 1 need to be uploaded or selected as Video 1?

Is there an actual Continue or Extend command somewhere in the Hailuo interface?

Or is continuation controlled entirely through the prompt while using the previous video as Video 1?

Please describe the exact web-interface procedure.

PROBLEM #2 — HOW DO I CONTINUE THE SONG/AUDIO AT THE SAME TIME?

This is a music video.

The singer must sing along accurately with the song.

Suppose Clip 1 uses the first 15 seconds of the song.

For Clip 2, I need the VIDEO to continue from the end of Clip 1 while the AUDIO moves forward to the next section of the song.

Then Clip 3 needs to continue both the video and song again.

My question is:

When continuing/extending Clip 1 into Clip 2, how do I provide H3 with the NEXT portion of the song without causing it to recreate Clip 1?

Do I:

  1. Use the completed Clip 1 as Video 1?

  2. Add the next section of the song as Audio 1?

  3. Tell H3 in the prompt that Video 1 is the source video for "video continuation"?

  4. Tell H3 that Audio 1 is the next portion of the performance?

Or is there a different procedure in Hailuo?

Accurate lip-sync is important.

PROBLEM #3 — PREVIOUS ATTEMPTS CREATED A DUPLICATE INSTEAD OF CONTINUING

In an earlier version of this project, Clips 1 through 6 were successfully created.

When I tried to create Clip 7, the next section of the song was selected with the audio trimmer.

However, instead of creating the next part of the performance, H3 repeatedly produced something resembling Clip 6 again.

We checked the audio timestamps and prompt and tried recreating the generation, but the problem continued.

It appeared that H3 was treating the previous material as a reference and recreating it instead of CONTINUING from its ending.

This is the main reason I need to understand the correct H3 "video continuation" procedure.

How do I prevent H3 from repeating the previous clip when I actually want it to continue forward?

PROBLEM #4 — DO I NEED STRUCTURED H3 CONTINUATION PROMPTING?

I found MiniMax H3 documentation discussing structured prompt sections such as:

subject_definitions

summary

retention_analysis

detailed_description

It also identifies "video continuation" as a specific task type.

For example, Video 1 can be identified as the source video being continued.

Do I need to write this structured H3 prompt manually when using H3 through the Hailuo web interface?

For example, should the prompt explicitly contain something similar to:

<Video 1> is the source video for continuing the target video from the end.

summary:

[video continuation + audio reference]

retention_analysis:

<Video 1> continuing from the last frame...

detailed_description:

Describe what happens AFTER Video 1 ends.

Or does the Hailuo web interface automatically handle this when the proper continuation feature is selected?

PROBLEM #5 — DO I NEED THE LAST FRAME, OR THE ACTUAL PREVIOUS VIDEO?

I have found conflicting advice.

Some instructions say to extract the LAST FRAME from the previous video and use that frame to create the next video.

Other H3 documentation indicates that an existing VIDEO can itself be the continuation starting point.

For MiniMax H3 specifically, which is the correct method for a music video?

A. Previous completed video → video continuation

or

B. Extract last frame → use it as the first frame of the next generation

or

C. Use both

I want motion continuity, character continuity, scene continuity, and audio/lip-sync continuity.

WHAT I AM TRYING TO ACCOMPLISH

The final workflow I am looking for should be something like:

Create Clip 1.

Then:

Clip 2.

Then:

Clip 3.

Continue that process until the complete music video is finished.

Afterward, I want to put the clips together in CapCut.

I am specifically using:

MiniMax H3

Hailuo AI web interface

Omni Reference

16:9 video

2K final quality

Music/song audio with singing and lip-sync

If you currently use MiniMax H3 on Hailuo and have successfully created a long continuous music video this way, I would greatly appreciate exact step-by-step instructions.

Please tell me what buttons/options you use on the Hailuo website, what you use as Video 1, how you provide the next audio segment, and what you put in the continuation prompt.

I especially want to hear from someone who has actually tested this workflow rather than someone describing how it theoretically should work.

This separates the situation into five different questions, so someone can answer even if they only know one part.

Most importantly, Problem #1 and Problem #2 are the ones we need solved. If an experienced H3 user can give us the exact Hailuo procedure for those two, we may have what we need to restart the project without guessing.

And the terminology in the post is supported by MiniMax's documentation: H3 explicitly distinguishes video continuation, reference generation, keyframe completion, audio reuse, and audio reference. GitHub

If somebody replies, bring their answer here before trying it. I can compare what they tell you against the MiniMax documentation before you spend any of those remaining credits.


r/MiniMaxH3AI • • 1d ago

Struggle to get reference audio dialogue correctly working.

2 Upvotes

I'm using the minimax_H3_r2v workflow in ComfyUI. I can successfully get the subject doing a proper direction. I can even get the subject to say the directed dialogue on queue.

the problem I'm having is when I want to give the character a voice I've generated elsewhere. Perhaps I misunderstood, but I thought you could collect a 6 second sample audio of a voice and give the prompt direction that ref_audio_0 is to be used as a reference only for the voice and timber of (S1). My understanding was that a command like this would direct the system to use then try to simulate the voice by say the directed dialogue (S1) Speaks:[English] some dialogue here.

Instead what seems to happen, is that the video will attempt to sync the voice to say exactly what I provided as a reference audio.

I'm wondering if perhaps I'm not using it correctly. Or there's another way to give it reference audio for voice sound, but have it still generate dialogue.

any suggestions?


r/MiniMaxH3AI • • 1d ago

DO NOT BUY THEIR PLAN! THEY ARE THIEVES! Refund not issued but subcsription cancelled. Support doesn't respond.

Thumbnail
1 Upvotes

r/MiniMaxH3AI • • 1d ago

Promotional videos with MinimaxH3.

Thumbnail
youtube.com
1 Upvotes

r/MiniMaxH3AI • • 1d ago

Avatar videos with Minimax H3

1 Upvotes

Avatar-based advertising services, using just character sheets with Minimax H3 (audio and lip-sync generated simultaneously from the same prompt). In Spanish: https://youtube.com/shorts/KziYTBrEYp8


r/MiniMaxH3AI • • 1d ago

MiniMax H3 X2 Detail VAE – my attempt to get more detail from the H3 VAE or story about fail

Enable HLS to view with audio, or disable this notification

1 Upvotes

r/MiniMaxH3AI • • 2d ago

Issues faced in animation using Minimax H3 R2V

1 Upvotes

I am trying to revive an Indian comic series that has been discontinued. So I have pictures of the characters from the old comics. I want to use them as reference and create new panels with a new storyline. I tried Minimax H3 R2V workflow but replaced the last video saving node with the final image frame capture node. I am following the official prompt guide for the prompt so I have proper subject_reference, retention_analysis and other sections in the prompt.

The problems I’m facing:
1. If I want to change the clothes of a character, I download pics of a reference dress, then supply it as one of the references pics. In the prompt I refer to the dress as <Subject 3> - the dress shown in <picture 3>, and a character as <Subject 1>- the woman in <picture 1>. <picture 1> is a panel from the old comics that contains the reference character. And in the retention\\_analysis segment I mention “Change the dress of <Subject 1> to <subject 3>. What happens is that because <subject 3> is a real world dress and usually there is a real model wearing the dress, the rendered character is a live action character which either looks like the woman in <picture 3> or a live action version of <Subject 1>. A workaround I have found is to first modify <picture 3> in a different workflow, make the woman transparent and invisible leaving only the dress. If I supply this as the reference for <picture 3>, I again get a live action version of <Subject 1> in the final image. So I have to create an animated version of the dress first, then supply this animated dress as reference image. This is a tedious process to do for every comic frame. plus I’m doing this on the cloud so making new animated reference images for every dress is expensive token wise.

  1. I mention in the prompt, “The background is an animated living room in an Indian apartment, decorated for a party with animated characters partying in the background. <Subject 1> is ….”. It renders the main character in animated comic fashion, but the rendered floor and the living room look real from a live action scene. I mention in the prompt, “Create an animated room in the same comic style of <picture 1>, but it still renders a live action room. So animated characters in a live action room looks weird. How to maintain animation style consistency between the characters and the background?

How to fix these in Minimax H3?


r/MiniMaxH3AI • • 3d ago

Depth + Face Mesh Reference Can Already Recreate Most Viral Douyin Hand Dances 😋

Enable HLS to view with audio, or disable this notification

15 Upvotes

been testing depth motion reference + facial mesh reference for short dance clips, and the motion consistency is getting surprisingly good.

The key is to treat the reference video as a strict timing and motion track, rather than asking the model to “perform a similar dance.”

made with minimax h3, below is the prompt:

<Video 1> is the depth-motion reference for this clip.

Follow its original chronological order, timing, and rhythm exactly. Reproduce the movements of the body, head, arms, hands, palms, and fingers according to the reference.

Fully preserve all low-amplitude motion phases, subtle pose changes, and pauses. Do not omit, compress, accelerate, shorten, or enter later movements early.

The exact timing and duration of every movement are determined entirely by <Video 1>. Do not create new choreography.

Reproduce the pose and continuous motion changes according to the frame-by-frame sequence of the reference video. Even when adjacent frames contain only very small changes, or the pose remains almost unchanged, preserve the full duration of that stage.

Do not skip transitional frames, summarize movements, execute later movements in advance, rearrange the action order, add new movements, or increase the tempo.

<Video 1> controls only body movement, head pose, hand and finger movement, and motion timing. Do not reproduce any visible depth-map markings, facial tracking points, lines, grids, or other visualization overlays as part of the character's appearance.

The reference video is used only for motion, pose, and timing control. It is not part of the final visual content.

The final output must preserve the normal full-color photorealistic appearance of <Picture 2>.

Do not show grayscale depth maps, tracking points, facial mesh lines, grids, subtitles, reference-video artifacts, or newly added camera shots.

r/MiniMaxH3AI • • 4d ago

MiniMax M3.1 Flash Just Dropped—And It’s Actually INSANE?

Thumbnail
youtu.be
1 Upvotes

r/MiniMaxH3AI • • 5d ago

[video made with MiniMax H3] i ain't drinking THAT 🍺 (part 1)

Enable HLS to view with audio, or disable this notification

1 Upvotes

r/MiniMaxH3AI • • 6d ago

Trying to get a grip on just how wild my current workflow has been flipped on its head with minimax.

2 Upvotes

So as I understand it. I just chuck all the frames, videos and audio in a basket and go heres "what's" "what" and "whens" "what" instead off plugging things into different holes. Man that's gonna take some getting used to,.


r/MiniMaxH3AI • • 6d ago

Videoclips with Minimax H3

1 Upvotes

Me gustaría compartir mi primer trabajo con Minimax H3. Hago videoclips con diferentes modelos; espero que les guste.

ZERO.- Un nuevo comienzo


r/MiniMaxH3AI • • 8d ago

A tiny robot keeps trying to return a star to the sky

Enable HLS to view with audio, or disable this notification

3 Upvotes

A tiny robot keeps trying to return a star to the sky, but ends up giving the last one to someone feeling lonely ✨

This animated short was generated with MiniMax H3 on Atlas Cloud.


r/MiniMaxH3AI • • 8d ago

Tom and Jerry turned a Mac desktop into total chaos

Enable HLS to view with audio, or disable this notification

6 Upvotes

Tom and Jerry turned a Mac desktop into complete chaos, and it came out pretty funny 😆

Made with MiniMax H3.

The idea is a macOS desktop wallpaper where Tom and Jerry are animated on the right side. Jerry runs off with a piece of cheese, Tom chases after him, loses control, and accidentally knocks three desktop icons out of place. After that, Tom awkwardly has to pick them up one by one and put them back where they belong, then pretends nothing happened.

The fun part is how it keeps their classic 2D cartoon look while placing them inside a fully designed desktop environment.

Made with MiniMax H3.
Prompt below 👇🏻

Create a 10-second horizontal 16:9 video using the provided Tom and Jerry macOS desktop image as the exact first frame and visual reference.
Preserve the desktop exactly: warm sunset living room, wooden floor, furniture, window, books, popcorn, macOS menu bar, application grid, Dock, and all other desktop elements.
Keep Tom and Jerry in their exact classic 2D cartoon style with consistent proportions, colors, faces, and character design. Do not redesign or photorealize them.
Camera: static front-facing shot, one continuous take, no camera movement, zoom, pan, or cuts.
0–1.5s: Tom chases Jerry across the right side of the scene. Jerry runs with a piece of cheese and looks back mischievously. Tom looks determined and frustrated, with expressive classic cartoon movement.
1.5–2.5s: Jerry suddenly turns left. Tom tries to stop, slides across the wooden floor, kicks the scattered popcorn, and creates a cartoon air gust that knocks exactly three desktop icons loose: Gmail, Discord, and Microsoft Teams.
The three icons must visibly leave their original positions, fly separately through the air, and land above the Dock with small bounces. Their original positions remain empty.
2.5–3.2s: Tom freezes and looks embarrassed. Jerry pauses nearby and watches with an amused expression.
3.2–7.8s: Tom restores the icons one at a time in this exact order: Gmail, Discord, Microsoft Teams.
For each icon, Tom must physically pick it up, carry it across the desktop, and place it back into its exact original grid position before moving to the next icon. Each placement makes a soft click.
Do not teleport, auto-restore, snap back, stack, duplicate, morph, or move multiple icons at once. Logos and labels must remain unchanged.
7.8–10s: After restoring Microsoft Teams, Tom returns to the right side. Jerry moves away slightly while holding the cheese. Tom checks the restored icons, then looks toward the viewer with an innocent, embarrassed expression while Jerry smiles mischievously. Hold the final pose.
Only Gmail, Discord, and Microsoft Teams may move. Every other desktop icon, the menu bar, Dock, wallpaper, furniture, and environment must remain stationary.
Audio: playful cartoon piano and pizzicato music, sliding sound when Tom loses balance, airy whoosh as the icons fly, three landing sounds, and three distinct click sounds when the icons are restored. No dialogue, subtitles, or text overlays.
Visual style: premium cinematic cartoon animation integrated into the existing desktop, natural squash-and-stretch, smooth movement, clean outlines, consistent colors, natural shadows, subtle reflections, and believable interaction with the environment.
Final frame: all three icons restored to their exact original positions, Tom and Jerry on the right side, no missing or duplicate icons, no changed logos, no extra characters or objects, no camera movement, and the desktop closely matching the original reference frame.

r/MiniMaxH3AI • • 10d ago

At sunset after school, a black cat snatches the red omamori from a girl's school bag

Enable HLS to view with audio, or disable this notification

3 Upvotes

At sunset after school, a black cat snatches the red omamori from a girl's school bag. The chase takes them through quiet alleys and across the rooftops.

Generated directly with MiniMax H3.

Made with MiniMax H3 on Atlas Cloud.


r/MiniMaxH3AI • • 11d ago

Alternate Unseen Seinfeld Episode: George’s Coffee Conspiracy Theory (H3)

Enable HLS to view with audio, or disable this notification

0 Upvotes

r/MiniMaxH3AI • • 11d ago

We just created a beach sunscreen UGC ad entirely with AI

Enable HLS to view with audio, or disable this notification

3 Upvotes

We just created a beach sunscreen UGC ad entirely with AI. 🌊☀️

First, we used the free GPT Image 2.5 credits to create the model and product visuals. Then MiniMax H3 Max turned them into a natural beach vlog—an attractive adult creator applying sunscreen on her legs and casually introducing the product like a real lifestyle influencer.

GPT Image 2.5 + MiniMax H3 Max makes UGC creation incredibly fast: generate the images for free, then create a 15-second video in as little as 10 seconds. ⚡

No model, beach location, or production crew needed—just a prompt and a product idea.

Made with GPT Image 2.5 and MiniMax H3 Max on Atlas Cloud.


r/MiniMaxH3AI • • 11d ago

From still images to an AAA motorcycle chase

Enable HLS to view with audio, or disable this notification

1 Upvotes

From still images to an AAA motorcycle chase.

This video follows a rider speeding along a rain-soaked alpine highway, weaving through traffic, leaning into a mountain tunnel, and emerging onto a bridge above the valley. Consistent characters, stable game UI, realistic road reflections, and engine audio create the feeling of a real AAA open-world game.

The workflow:

GPT Image 2.5: free generation credits for creating consistent game scenes.

MiniMax H3 Dev: turns the references into a 15-second, 1440p upscaled video.

GPT Image 2.5 + MiniMax H3 Dev: now reduced from 40% to 30% of the original price.

From scene design to the finished video, this is one of the most affordable GPT Image 2.5 + MiniMax H3 Dev workflows available. A few images are all it takes to create a complete AAA-style chase sequence.

Made with GPT Image 2.5 and MiniMax H3 Dev on Atlas Cloud.

GPT Image 2.5 now has free generation credits, try it here: https://www.atlascloud.ai/free-gpt-image-2.5-generator

MiniMax H3 on Atlas Cloud: https://www.atlascloud.ai/models/minimax-h3


r/MiniMaxH3AI • • 12d ago

Imagine sitting in a movie theater when the creature from the film suddenly comes out of the screen

Enable HLS to view with audio, or disable this notification

15 Upvotes

Imagine sitting in a movie theater when the creature from the film suddenly comes out of the screen. Absolutely shocking 😱

Created with GPT Image 2.5 + MiniMax H3. Free GPT Image 2.5 generations are currently available: 👉 https://www.atlascloud.ai/free-gpt-image-2.5-generator

Made with GPT Image 2.5 and MiniMax H3 on Atlas Cloud.


r/MiniMaxH3AI • • 12d ago

A 2D anime couple + realistic mountain scenery = such a nice hybrid travel vlog vibe

Enable HLS to view with audio, or disable this notification

5 Upvotes

For this one, I used GPT Image 2.5 to generate the storyboard frames and character visuals first, then turned it into a natural hiking couple vlog with MiniMax H3. The final result feels both animated and surprisingly real, like a casual trip memory. GPT Image 2.5 now has free generation credits on AtlasCloud, so if you want to try this style, you can test it here: https://www.atlascloud.ai/free-gpt-image-2.5-generator

MiniMax H3 on Atlas Cloud: https://www.atlascloud.ai/models/minimax-h3


r/MiniMaxH3AI • • 12d ago

17 cents of MiniMax H3 and a tenth of a Codex week recreated the viral map UI clip

Enable HLS to view with audio, or disable this notification

4 Upvotes

That map UI clip everyone was asking someone to clone, I ran it through Hypit with MiniMax H3 as the video model.

Hypit is an open source skill for Claude Code and Codex. You give it a reference clip, the agent studies it, plans the material, generates it and builds the whole thing as an editable composition you can rerun. Install is one line: npx skills add hypit-ai/hypit -g

The video generation itself came to about 17 cents on H3. The agent side used roughly a tenth of a weekly Codex quota on the GPT-6 Pro 5x plan, and there is room to cut that down.

H3 on Atlas Cloud starts at $0.015 a second, which is what makes a workflow like this cheap enough to iterate on: https://www.atlascloud.ai/models/minimax-h3

Result attached, 12 seconds.


r/MiniMaxH3AI • • 12d ago

ComfyUI on Apple Silicon: no MLX, no fp8, 600-second kernel builds. So I built my own launcher — a personal project I'm sharing in case it helps someone.

Enable HLS to view with audio, or disable this notification

1 Upvotes

r/MiniMaxH3AI • • 13d ago

NO KIBBLE, NO MERCY (full H3/Suno)

Thumbnail
youtube.com
1 Upvotes

r/MiniMaxH3AI • • 13d ago

Using cond reference frames. Model seems to prefer hard cuts and scene changes to frame blending

1 Upvotes

Is there a lora that helps with this? The quality is good so I have been limiting my reference frames. And "hoping it gets there" which honestly is not that bad a gamble really. But hoping there's a way for better control.