Files
comfyui-workflows/optimized/notes/minimax.md
T

46 lines
2.3 KiB
Markdown

# MiniMax Notes
## Current default assumptions
- `MiniMax H3` is promising, but brute-force paths are too heavy to be the first bet on a `3070`.
- Speed-first and low-VRAM community workflows matter more than official purity for the first real test.
## Strong current source workflows
- `models/minimax/minimax-h3-official-t2v/`
- `models/minimax/minimax-h3-official-i2v/`
- `models/minimax/minimax-h3-turbo-lora-community/`
- `models/minimax/minimax-h3-int8-r2v-community/`
- `models/minimax/minimax-h3-extender-ref2va/`
## Strong current heuristics
- Prefer Turbo LoRA for the first `text-to-video` attempt.
- Prefer INT8-based community workflows for `reference-to-video` exploration.
- Keep attention and memory-efficiency helpers if they are already integrated and well documented.
## Current weak areas
- No proof yet that MiniMax is truly practical on Chris's box.
- Custom-node stack is more fragile than the current `flux2klein` path.
## Best current optimization direction
- Keep `optimized/text-to-video/minimax-h3-3070-turbo-master/` as the first MiniMax default.
- Keep the INT8 reference-video workflow as the next donor for future low-VRAM optimization.
- Keep `optimized/image-to-video/minimax-h3-extender-3070-r2v-480p-clip-by-clip/` as the first `MiniMax H3 Extender` path to compare against the compact `Contex Loop` masters.
## Ref2VA prompt guidance
- For `MiniMax H3 Ref2VA`, treat `480p` and `10s` as the safe planning baseline on `8GB VRAM` unless direct tests prove otherwise.
- Use one master-scene reference to control environment/composition, one primary full-body reference per subject to control overall appearance, and headshots only to reinforce facial identity.
- Explicitly state when multiple images represent the same subject, and explicitly tell the model to ignore character-reference backgrounds so they do not fight the master environment.
- Use the prompt mainly for action, performance, camera movement, dialogue, and audio rather than re-describing scene details already visible in the references.
- Full guidance now lives in `optimized/notes/minimax-h3-ref2va-prompt-guidelines.md`.
## What needs Monday proof
- whether the Turbo LoRA workflow is just technically possible or actually usable
- whether the source-recommended `8`-step path holds up in quality
- whether MiniMax feels like a real challenger to `LTX` locally