---
license: other
license_name: minimax-h3-community-license
license_link: https://huggingface.co/MiniMaxAI/MiniMax-H3/blob/main/LICENSE
pipeline_tag: video-to-video
tags:
- minimax-h3
- comfyui
- workflow
- ref2v
- prompt-enhancer
- character-replacement
- v2v
---
# Nugget H3 EasyR2V — Prompt Enhancer & Full Workflow
**Why spend 30 minutes doing something when you can spend 2.5 weeks making a tool to automate it.** That's basically the whole reason this exists.
I wanted an easy to use workflow for us simple folk who want to do **Ref2V** on MiniMax H3 but keep running into troubles with the model understanding me. A good prompt does 80–90% of the work with H3 R2V — so instead of writing them by hand every time, this workflow writes them for you.
Think of it as **"Ref2Video for dummies"**. But it's not just limited to that.
---
## What it does
- **Scans your video** (if you're using one) to caption it and transcribe the audio
- **Captions all your images** - so it also works as a pure image-to-video workflow
- **Loads a small LLM of your choice** and writes your H3 R2V prompt in the correct format with your stated intent (user prompt)
- **If you're on the Full workflow**, it generates the video too
## What it does NOT do
- **Be creative for you** - The current WF is only setup to do the prompt formatting, it is not able to generate new ideas for you
- **It cannot perform magic** - You are still limited to what the H3 model can and cannot do. Complex scenes are still very difficult
### Saving Time
- If your video doesn't change, a second run doesn't trigger a new video transcription (If you have "FIXED SEED")
- If your images and user prompt don't change, it doesn't write a new prompt — so you can re-run on a new seed to test without paying the LLM cost every time (If you have "FIXED SEED")
---
## Two workflows
| File | What it does |
|---|---|
| `Nugget_H3_EasyR2V_Full_WF_v03.json` | The full thing — transcribe + prompt + generate video |
| `Nugget_H3_EasyR2V_Prompter_Only_v03.json` | Just writes the H3 prompt. Copy it out, or wire it into your own H3 graph |
Prompter-only is handy if you've already got your own H3 setup dialled in and just want a better prompt without ripping your workflow apart.
**Full workflow:**

**Prompter only:**

---
## Examples
**Character replacement** — swap the person in a clip for one from your reference images:

Side by side, original vs replaced:
**Background + character replacement** — swap both in one go:
**Voice line + character replacement** — new person, new dialogue:
---
## Setup
Both workflow files have a big **START HERE — Downloads & Setup** markdown note pinned inside them, with every model link, folder path and install step. Rather than duplicate it all here, just open the workflow and read the note — it's more accurate than a copy of it would be.
The short version:
- **Custom nodes**: ComfyUI-Nugget, ComfyUI-KJNodes, and (Full workflow only) ComfyUI-PlagueKind-Nodes
- **Run the Nugget install script** if you want dialogue transcribed — it installs `faster-whisper` into ComfyUI's own Python. A normal `pip install` goes into the wrong interpreter and it will still say the package is missing
- **H3 models** from [🤗 Comfy-Org/MiniMax-H3](https://huggingface.co/Comfy-Org/MiniMax-H3)
- **Prompt enhancer model** — start with Qwen3-VL 8B fp8_scaled. Drop to nvfp4 on 12 GB cards
If the file won't open in ComfyUI, a node pack is missing. Install it, restart, reopen. Bypassing won't help.
---
## Tips for good character replacement
- **Don't expect miracles.** - It is still H3 model and sometimes tempermental. Check your enhanced prompt and consider rolling again if it is not right.
- **H3 is a tool, you're the one using it.** If you don't specify emotions, expect expressionless results. This current setup will only do what you intend for it to do. **Slop prompt in, slop video out**
- **Limit the resolution and length.** There seems to be an arbitrary context window that may be linked to your specs, so the higher higher res/longer time your input/outputs. Keep it smaller and your success rate goes up
- **If the video is easy, replacement should be easy too.** H3 has a quirk though — if the original person and the new person look too similar, it sometimes converges back to the original. A prompt won't always fix that. If you hit it, consider changing the person to a intermediate step (faceless green person). The new body/face will transfer over better. Alternatively you can look into Sam3 character replacement method.
- **More than one person in the scene?** Describe the scene properly. `replace the man wearing white shorts with the man in ` beats `replace the man with ` every time.
- **Complex scenes?** It will be very difficult (I've tried), scenes with too many people, too many cuts, characters obstructed are very difficult for the model to properly identify and swap.
- **Give the LLM some context.** A one-liner in the user prompt like `