Привет!
Представляю рабочий процесс для генерации видео на модели MiniMax H3 в ComfyUI.
Это не набор разрозненных нод, а собранная схема — от загрузки картинки до готового видео. Всё протестировано вручную, в разных режимах, на разном железе. Ничего не падает, ничего не вылетает по памяти.
Если есть замечания по работе процесса и нод, пишите.
Внутри используются кастомные ноды. Я писал их сам, потому что нормальных аналогов не нашёл — либо не хватало функционала, либо работали нестабильно.
Все ноды могут работать автономно, нода контроля камеры может работать без языковой модели в ручном режиме, выберете пресет и нода впишет поведение камеры в ваш промт.
нода загрузки изображения: мини Фотошоп, в реальном времени (до генерации) поддержка в превью до 2к (далее системные ограничения) но работает с любым разрешением до 8к, возможно подтянуть: яркость, контрастность и так далее. имеет пресеты на стандартные разрешения. Есть ручной режим, ставте любое разрешение. Имеет защиту кратности (по умолчанию 32, важно для моделей минимакс и ЛТХ, можно изменить). Нода полностью автономна.
Нода загрузки лор (всего 5) не оставляет после себя фантомов при отключении, память вычищается полностью, если лоры взаимоисключающие то порядок такой: первая(1) лора основная, исключающая или резко меняющая картинку должна быть пятой(5). это нужно для того чтобы нода подчистила хвосты и не сваливала все в кучу. В остальном порядок любой.
последняя нода исключает выгорание при малых шагах генерации, также она умеет сама подстраивается под вышу память на видеокарте, в том числе и апскейлер так же можно настроить (быстро-качественно). апскейлер так же подстраивается под вашу память на видеокарте и оперативную память, однако чудес не ждите, оперативная память при апскейле должна быть не менее 16 гигов..чем больше тем лучше. Модель апскейла любая на ваш вкус.
—
ЧТО ВНУТРИ
• Автопромптер — короткий текст превращается в технический промпт для MiniMax H3.
• Контроль камеры — модель больше не летает, куда вздумается. Вы управляете.
• Interactive Image Processor — мини-фотошоп прямо в ноде. Превью в реальном времени.
• Manual Color & Upscale Panel — цветокоррекция и AI-апскейл с обработкой порциями.
• До 5 слотов LoRA — каждый включается отдельно, со своей силой.
• RIFE — интерполяция кадров 24 → 48 fps.
• Режимы 8 / 20 шагов — турбо (быстро) или обычный (стабильно).
—
ТРЕБОВАНИЯ К ЖЕЛЕЗУ
Минимум: 8 ГБ VRAM / 16 ГБ RAM (проверено лично)
Стандарт: 12 ГБ VRAM / 32 ГБ RAM
Рекомендуется: 16 ГБ VRAM / 32 ГБ RAM
Скорость: 15 секунд, 1024×1024, без апскейла — около 6 минут (RTX 5070 Ti).
—
УСТАНОВКА
1. Скачайте архив с GitHub.
2. Распакуйте.
3. Скопируйте три папки в ComfyUI\custom_nodes\:
ComfyUI-H3-Lora-Selector
ComfyUI-H3-Static-Camera
ComfyUI-Interactive-Processor
4. Запустите ComfyUI заново.
5. Откройте воркфлоу из папки workflow/.
📥 Скачать: https://github.com/neron5480-cell/neron-video-minimax-h3-i2v/releases/latest
Полная инструкция (модели, внешние ноды, troubleshooting) — в README_INSTALL.txt в архиве.
—
ЧТО НОВОГО В v3.0
• Manual Color & Upscale Panel: обработка порциями, автоматическое уменьшение порции при нехватке VRAM, fp16, режим Fast.
• Interactive Image Processor: чёткое превью, скрытие неиспользуемых полей.
• Воркфлоу: апскейл до RIFE (вдвое меньше кадров), RIFE в отдельном сабграфе, очистка VRAM в цепочке.
• Контроль камеры: перехват промпта после автопромптера, вписывание координат без переписывания текста.
• 5 слотов LoRA: каждый включается отдельно, со своей силой.
—
ПРОМПТ — ПРОСТО
Автопромптер многоязычный (русский + английский). Пишите как думаете — хоть одной строкой. Он сам развернёт в технический промпт с камерой, таймингами и звуком.
Чем подробнее промпт — тем точнее результат. Но даже короткого хватит.
Пример:
"15 секунд видео. Девушка идёт домой, камера облетает её и застывает за левым плечом."
На выходе — полноценный технический промпт с углами, дистанцией, таймингом и звуком.
—
ЛИЦЕНЗИЯ
MIT. Используйте свободно.
—
АВТОР
Roman (neron5480-cell)
GitHub: https://github.com/neron5480-cell
Civitai: https://civitai.com/user/neron5480922
Если что-то не работает — пишите в комментариях. Я читаю.
—
ENGLISH
Hi!
This is a working workflow for generating video from an image using the MiniMax H3 model in ComfyUI.
It's not a pile of unrelated nodes — it's a complete pipeline from image upload to final video. Tested manually, in different modes, on different hardware. No crashes, no out-of-memory errors.
Custom nodes are included. I wrote them myself because I couldn't find any working alternatives.
All nodes can operate autonomously; the camera control node can function without a language model in manual mode—simply select a preset, and the node will incorporate the camera behavior into your prompt.
If you have any comments regarding the operation of the process and nodes, please let me know.
Image upload node: acts as a mini-Photoshop with real-time editing (prior to generation). Preview support up to 2K (limited by system constraints thereafter), though it handles resolutions up to 8K. Allows adjustments such as brightness, contrast, and more. Includes presets for standard resolutions and a manual mode for custom resolution settings. Features aspect ratio/dimension alignment protection (defaulting to a multiple of 32—important for Minimax and LTX models, but adjustable). The node is fully autonomous.
The LoRA loading node (supporting up to 5) leaves no phantoms behind upon disconnection; memory is cleared completely. If the LoRAs are mutually exclusive, the order matters: the first (1) LoRA should be the primary one, while any LoRA that is exclusive or drastically alters the image should be the fifth (5). This ensures the node cleans up any leftovers and avoids jumbling everything together; otherwise, the order does not matter.
The latest node eliminates "burnout" during small generation steps and automatically adjusts to your GPU's VRAM; the upscaler can also be configured for either speed or quality. The upscaler adapts to both your VRAM and system RAM, though you shouldn't expect miracles—you need at least 16 GB of system RAM for upscaling (and the more, the better).You can use any upscaling model you like.
—
WHAT'S INSIDE
• Autoprompter — short text turns into a technical MiniMax H3 prompt.
• Camera control — the model no longer flies wherever it wants. You decide.
• Interactive Image Processor — a mini-Photoshop right inside the node, with live preview.
• Manual Color & Upscale Panel — color correction and AI upscaling with chunked processing.
• Up to 5 LoRA slots — each one toggles independently with its own strength.
• RIFE — frame interpolation 24 → 48 fps.
• 8 / 20 step modes — turbo (fast) or normal (stable).
—
HARDWARE REQUIREMENTS
Minimum: 8 GB VRAM / 16 GB RAM (verified personally)
Standard: 12 GB VRAM / 32 GB RAM
Recommended: 16 GB VRAM / 32 GB RAM
Speed: 15 seconds, 1024×1024, no upscale — about 6 minutes (RTX 5070 Ti).
—
INSTALLATION
1. Download the archive from GitHub.
2. Unpack it.
3. Copy the three folders into ComfyUI\custom_nodes\:
ComfyUI-H3-Lora-Selector
ComfyUI-H3-Static-Camera
ComfyUI-Interactive-Processor
4. Start ComfyUI again.
5. Open the workflow from the workflow/ folder.
📥 Download: https://github.com/neron5480-cell/neron-video-minimax-h3-i2v/releases/latest
Full instructions (models, external nodes, troubleshooting) — in README_INSTALL.txt inside the archive.
—
WHAT'S NEW IN v3.0
• Manual Color & Upscale Panel: chunked processing, automatic chunk reduction on low VRAM, fp16, Fast mode.
• Interactive Image Processor: sharp preview, unused fields hidden.
• Workflow: upscale before RIFE (half the frames), RIFE in its own subgraph, VRAM cleanup in the chain.
• Camera control: intercepts the prompt after the autoprompter, writes coordinates without rewriting your text.
• 5 LoRA slots: each one toggles independently with its own strength.
—
PROMPT — KEEP IT SIMPLE
The autoprompter is multilingual (Russian + English). Write however you think — even one line. It will expand into a technical prompt with camera, timing and sound.
The more detail — the more accurate the result. But even a short one works.
Example:
"15 seconds of video. A girl walks home, camera orbits her and locks behind her left shoulder."
Output: a full technical prompt with angles, distance, timing and sound.
—
LICENSE
MIT. Use freely.
—
AUTHOR
Roman (neron5480-cell)
GitHub: https://github.com/neron5480-cell
Civitai: https://civitai.com/user/neron5480922
If something doesn't work — leave a comment. I read them.


