r/comfyui 9h ago

Tutorial [Tutorial] Create AI Anime Videos Locally with ComfyUI + MiniMax H3

Enable HLS to view with audio, or disable this notification

101 Upvotes

In this tutorial, I show you how to create a 90s anime-inspired AI video completely locally on your PC using ComfyUI, MiniMax H3, and Krea 2 Turbo.

I walk through the full workflow from reference image generation to final video generation.

You’ll learn how to:

  • Generate consistent anime reference images using the Krea 2 Turbo text-to-image workflow
  • Create separate references for the character, environment, vehicle, and props
  • Use multiple reference images with the MiniMax H3 Reference-to-Video workflow in ComfyUI
  • Structure prompts so MiniMax H3 understands which reference image controls each part of the scene
  • Describe camera framing, character actions, object placement, animation, lighting, and timing
  • Create a 90s hand-drawn anime look with lower-frame-rate animation
  • Run the entire workflow locally without monthly AI video subscriptions

In the example, I use separate reference images for the character, convenience store environment, car, and skateboard, then combine them into a single animated anime scene.

I also explain how I approach prompting for reference-to-video generation, including reference assignment, shot description, motion instructions, camera constraints, and visual consistency.

Workflow, prompts and reference files

https://drive.google.com/drive/folders/1qI9Oi5gqWJh8xIwHYsn0S25DPqK8M0XI?usp=sharing

MiniMax H3 models

https://docs.comfy.org/tutorials/video/minimax/minimax-h3

Krea 2 Turbo T2I models

https://comfyui.org/en/krea-2-open-source-models-are-now#content-required-models


r/comfyui 12h ago

News A quick Minimax H3 news round-up - 1st September 2026

65 Upvotes

Another quick Minimax H3 news and goodies round-up, for those who may have missed some items.

-> BUNNY General Motion Continuity Repair LoRA. This is said to be the best H3 motion-fixer LoRA, by those have tested several such helpers. "A general-purpose motion support LoRA, not a combat-only LoRA. It can be used for running / sports / dance / acrobatics / character interaction / combat / weapon motion and other dynamic scenes. ... designed around repairing those ['awkward movement'] moments rather than simply making every motion stronger." The trigger is bunny_crisp_motion but it will work without a trigger in the prompt. No workflow.

https://huggingface.co/JOKER141/MiniMax-H3-General-Motion-Continuity-Repair

-> MiniMax-H3-Combat-Base-V2 LoRA. "Combat Base V2 is no longer just a combat LoRA — it now works for both action and dialogue scenes, adding richer body movement, finer sound details, and stronger shot-to-shot continuity to character performance." No trigger needed, unless you have really high intensity (e.g. Marvel superheroes battle) where you use prfight2 or "heavy stunt impact" (e.g. Indiana Jones style) where prfin1 is the trigger. Has two demo workflows, for FL2VA and Ref2VA.

https://huggingface.co/JOKER141/MiniMax-H3-Combat-Base-V2

-> MiniMax H3 Prompt Queue. This ComfyUI custom node for FL2VA... "provides a practical Fixed Prompt + Motion Prompts workflow so you can store many motion prompts in one node and queue them sequentially." Has a workflow.

https://github.com/AlixAsset/minimax-h3-prompt-queue

-> Batch Minimax for ComfyUI. Appears to run one reference image against one video reference, along with one prompt file (e.g. video_1.mp4 + refimage_1.png + prompt_1.txt), and it then runs the next set. Has workflows.

https://github.com/TagirovAlex/BatchMiniMax

-> Prompts, assets and workflows for the Minimax H3 short film "The Mole Beyond The Stars".

https://github.com/GeeKanJi/MiniMax-H3---Workflow-for-The-Mole-Beyond-the-Stars

https://www.youtube.com/watch?v=kX6UmSp5r-g

-> Workflow and assets for a pilot episode for "Quibble", in which a would-be super-villain gets a parking ticket. The maker is focused on locking down the appearance of a complex stylised toon character, while still making him able to act.

https://github.com/mkhamra/quibble-h3

-> ComfyUI-MiniMaxH3-CLSS. "Port of the LTX-2.3 CLSS package to the H3 architecture". Said to prevent the gradual scene/character collapse that can happen when generating 30-second+ videos with Minimax clip-chainers. Appears to requires a graphics card with 16Gb of VRAM.

https://github.com/nazgut/ComfyUI-MiniMaxH3-CLSS

-> New lyrics-focused LoRA sliders for use with Minimax Music. Includes sliders for 'fierce', 'grit', 'sexy', and 'joy', among others.

https://huggingface.co/ntc-ai/minimax-music3-concept-sliders

-> And finally, the excellent open-source video editor OpenShot has just released version 4.0. Among the new features are improved export presets, plus... "creative presets for motion, camera moves, colour, film looks, lighting, and audio", and... "local AI model downloads and workflows for Object Mask and Object Detection". Vital tip for first-time users: To zoom into the timeline, hold down the Crtl key and then roll the mouse-wheel forwards. OpenShot also officially provide some experimental ComfyUI integration nodes and workflows.

https://github.com/OpenShot/openshot-qt/releases/tag/v4.0.0 (desktop versions and changelog)

https://github.com/OpenShot/OpenShot-ComfyUI

~ OLD POSTS ~

https://old.reddit.com/r/comfyui/comments/1w3rgh6/a_quick_minimax_h3_news_roundup_31st_august_2026/

https://old.reddit.com/r/comfyui/comments/1w2mfd9/a_quick_minimax_h3_news_roundup_30th_august_2026/

https://old.reddit.com/r/comfyui/comments/1w1wpkz/a_quick_minimax_h3_news_roundup_29th_august_2026/

https://old.reddit.com/r/comfyui/comments/1w11jri/a_quick_minimax_h3_news_roundup_28th_august_2026/

https://old.reddit.com/r/comfyui/comments/1w053ce/a_quick_minimax_h3_news_roundup_27th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vyzvqi/a_quick_minimax_h3_news_roundup_26th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vy4drz/a_quick_minimax_h3_news_roundup_25th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vx0duv/a_quick_minimax_h3_news_roundup_24th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vwl8do/a_quick_minimax_h3_news_roundup_23rd_august_2026/

https://old.reddit.com/r/comfyui/comments/1vvkmra/a_quick_minimax_h3_news_roundup_21st_august_2026/

https://old.reddit.com/r/comfyui/comments/1vuihag/a_quick_minimax_h3_news_roundup_21st_august_2026/

https://old.reddit.com/r/comfyui/comments/1vtgs7b/a_quick_minimax_h3_news_roundup_20th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vsjzrp/a_quick_minimax_h3_news_roundup_19th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vrsspo/a_quick_minimax_h3_news_roundup_18th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vqyn8p/a_quick_minimax_h3_news_roundup_17th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vq5d5u/a_quick_minimax_h3_news_roundup_16th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vpbtx2/a_quick_minimax_news_roundup_15th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vojtjd/a_quick_minimax_news_roundup_14th_august_2026/


r/comfyui 5h ago

Tutorial MiniMax H3 Just Got MUCH Better — Cleaner Video, Better Faces & 1080p Workflow

Enable HLS to view with audio, or disable this notification

8 Upvotes

Just came across a kind of weird lora combination that can fixthe oily/plastic look from acceleration LoRAs.

I also added a two-stage latent upscale workflow, to fix blurry distant faces, and quality loss in high-motion scenes. You can also generate a quick low-res preview first, find a good seed, and then refine it to around 1080p. This saves quite a bit of time compared to doing the full-resolution generation every time.

The workflow also uses a merged H3 model, so text-to-video, image-to-video, first/last frame, and multi-reference generation can all be handled in basically the same workflow.
Tutorial (Workflow Included): https://youtu.be/RicFavgpL5o


r/comfyui 11h ago

News MiniMax H3 acceleration arena/leaderbord: 15+ H3 LoRAs, fine-tunes, Max

Thumbnail
huggingface.co
17 Upvotes

r/comfyui 17h ago

Workflow Included testing t2va fast h3, only 1 minute generating each video on 0.4mp resolution

Enable HLS to view with audio, or disable this notification

47 Upvotes

1 minute on 0.4 mp resolution
2 minute on 0.5 mp resolution
using RTX 4060Ti 16GB VRAM
workflow: https://civitai.com/models/2906467/fast-minimax-h3-t2va?modelVersionId=3286956


r/comfyui 3h ago

Help Needed Question About MiniMAX H3

2 Upvotes

I've been trying out MiniMAX H3, and overall I think it's really great. The one thing that keeps bugging me, though, is that consistency falls apart when I generate a video from a photo. For example, if I feed in a AI person's photo to create a talking scene, the person's face starts drifting the moment any movement kicks in — it keeps the general vibe of the original but basically morphs into a completely different person. I've also tested LTX, and oddly enough, MiniMAX H3 seems to struggle more with maintaining facial consistency than LTX does. I'm still pretty new to AI generation, so there's a lot I don't fully understand yet. If anyone has tips or workarounds for this, I'd really appreciate the help!


r/comfyui 6h ago

Resource I vibe coded a gallery extension for ComfyUI so you can browse outputs and reload the exact workflow that made them

Thumbnail
3 Upvotes

r/comfyui 10h ago

Resource A radial menu for ComfyUI: 😺 NKD Radial Menu

Enable HLS to view with audio, or disable this notification

6 Upvotes

I built a radial menu for ComfyUI designed to work with gestures and pure muscle memory, just like I’ve been dreaming of since day one working with Maya.

It also hooks into my NKD Reroutes extension, which lets you snap nodes by proximity, align them with their connections, bridge connections remotely, and clean up the entire workflow with a single click.

Fully customizable, of course. You can throw in whatever nodes you want and tweak it to your liking with custom icons, colors, and categories. But since I know most people won't bother setting it up manually, I wired up an MCP server so you can hook it up to any AI agent and have it build your custom setup just talking with Jarvis.

https://github.com/Nekodificador/ComfyUI-NKD-Radial-Menu


r/comfyui 1h ago

Workflow Included Cursor

Upvotes

Hello , can anyone guide me how to use cursor Ai, i have been using this for a week but it looks little bit complicated for me


r/comfyui 1d ago

Resource Finally I can see a truly huge jump in speed

135 Upvotes

The minimax_h3_fl2v_turbo_4step_v0.1_768p_sla_comfyui_bf16.safetensors lora combined with the H3 SLA attention node has chopped my speeds down so much that i've started rendering higher. I'm doing 1920x1088 10 second clips in 5 mins for i2v and 154 seconds in t2v.

5090, 64gb ram.

I can go back and see how fast it was doing 1280x768 later if anyone cares. It was blistering though. That node has made an epic improvement. I mean I wanna say half, but I kinda feel like it's more than that. I wasn't even rendering 1920x1088 bc it took too long and 1280 was fine. But tbh, 1920x1088 gives much more detail, and I know you can upscale, but idk, I think it's more than that.

Anyway worth checking out.

https://github.com/PlagueKind/ComfyUI-PlagueKind-Nodes/tree/main/ComfyUI-H3-SLA-Attention

https://huggingface.co/lightx2v/Minimax-h3-Turbo-SLA/blob/main/minimax_h3_fl2v_turbo_4step_v0.1_768p_sla_comfyui_bf16.safetensors


r/comfyui 2h ago

Help Needed Beginner guidance needed: Face-swapping, photorealism (Indian skin tones), & workflow suggestions for 4GB VRAM (RTX 3050 Laptop)

0 Upvotes

Hey everyone! 👋

I'm looking to transition from prompt-based online tools (using Gemini for concepts/prompts and Perchance) to building a local, controlled pipeline in ComfyUI.

Since I am running on a budget laptop, I want to construct a reliable, light workflow for generating photorealistic selfies, person-to-person face swaps (both SFW and uncensored/NSFW), with a specific focus on achieving authentic Indian skin tones, natural lighting, and realistic texture (avoiding that smooth, plastic AI look).

My Setup:

  • GPU: NVIDIA GeForce RTX 3050 Laptop (4GB VRAM)
  • System RAM: [Insert your RAM size here, e.g., 16GB / 32GB]
  • Launch flags: --lowvram --use-pytorch-cross-attention

Looking for Model & Workflow Suggestions:

  1. Recommended SD 1.5 Photorealism Checkpoints:
    • Since SDXL and FLUX are too heavy for 4GB VRAM, which SD 1.5 checkpoints yield the most lifelike skin micro-textures out of the box?
    • (Currently considering CyberRealistic v5.0, Realistic Vision v6.0, or Desi Tadka SD 1.5). Are there other underrated checkpoints for photorealism?
  2. LoRAs / Prompting for Authentic Indian Skin Tones:
    • What specific LoRAs or embeddings are best for generating accurate South Asian/Indian skin tones without making them look gray, washed out, or over-processed?
    • Has anyone tested the Real Indian Beauty LoRA or Desi Skin Tone Sliders on SD 1.5? How do they blend with realism checkpoints?
  3. Lightweight Face-Swap & Restoration Pipeline:
    • What is the most VRAM-efficient face-swap stack for ComfyUI in 2026?
    • Is ReActor (inswapper_128) still the best option for low-VRAM GPUs?
    • To fix the soft 128x128 output from inswapper_128, which restore model (CodeFormer vs. GPEN-BFR-512) provides the most natural skin texture without looking fake?
  4. Workflow Architecture Questions:
    • Would you recommend generating the base prompt/selfie background first and running a 2-pass / separate Face-Swap step, or chaining KSampler ➔ ReActor ➔ Face Restore into a single combined node graph?
    • Does anyone have a starter .json workflow template suited for 4GB VRAM that I could study or load directly?

I'd really appreciate any workflow screenshots, node suggestions, or .json templates that can help a beginner get started smoothly without hitting Out-Of-Memory (OOM) errors! Thanks! 🙏


r/comfyui 12h ago

Resource I created a node for posting to CivitAI with the models already embedded

Post image
7 Upvotes

Pretty straight forward, add this node into your workflow, connect the text prompt and final image/video.

Node resolves the models with SHA256 from the CivitAI API and links them automatically.
Requires civitai_token in your environment variables.

https://github.com/Hearmeman24/ComfyUI-CivitAI-Publisher

Drop a star if you find this useful


r/comfyui 2h ago

Help Needed I want to start using comfy comming from Automatic. Feeling lost.

1 Upvotes

HI there. Excuse the english. not my first language.

I stopped using image gen on automatic about 3 months ago and want to step into Comfy Ui super specificly Fn-Moment Anima-Turbo. I have no Idea how to use comfy ui. I've played with it a couple of months ago but got super lost. Is there any documentation videos as how to use it?

I would like to run a cloud instance. I have access to a cloud service providor that allows spicy anime stuff which is what I want but I would like to run something on runpod but I've seen mixed reviews.

I will not have access to run locally so not an option for me.

All help or pointers are appreciated!


r/comfyui 1d ago

News Trellis.2 and Pixal3D Are Now Native in ComfyUI

Thumbnail
gallery
117 Upvotes

Both Trellis.2 (Xiang et al., 2025) and Pixal3D (Li et al., 2026) now run natively in ComfyUI. No custom nodes, no compiled CUDA extensions, no PyTorch downgrades, and no non-commercial dependencies.

This is more than a model integration. It ships with a rebuilt 3D pipeline: new Load/Preview/Save 3D nodes, a set of mesh post-processing nodes, and an extended PBR texturing stage that bakes normal and ambient occlusion maps for a complete material set. Everything runs on consumer hardware, and everything is free to use, including commercially.

Why Trellis.2 still matters, ten months later

When Microsoft open-sourced Trellis.2 in December 2025, it immediately became the best open-source model for 3D generative AI. A 4-billion-parameter model built on a compact structured latent representation (O-Voxel). It generates high-fidelity 3D assets from a single image at effective resolutions up to 1536³, handling complex topologies that earlier methods struggled with. It also shipped with a PBR texturing model generating base color, roughness, and metallic maps.

Ten months is an eternity in generative AI, yet Trellis.2 hasn’t just aged well, it has become foundational. Several open-source 3D models released since build directly on it, the most notable being Pixal3D whose implementation uses the Trellis.2 backbone.

The community got there first

As always, the ComfyUI community was quick to bring Trellis.2 into the graph. Within days of the release, custom node packs appeared, the most popular being ComfyUI-TRELLIS2 by Andrea Pozzetti and ComfyUI-Trellis2 by VisualBruno, which together gathered well over a thousand stars. We’re grateful to both authors as they proved the demand and carried the community for months.

Despite their efforts, running Trellis.2 remained a challenge for two reasons.

Installation

The original implementation targets environments built around PyTorch 2.6.0 with CUDA 12.4, which for many users meant downgrading their existing ComfyUI environment. On top of that sit a stack of compiled CUDA extensions (flash-attention, FlexGEMM sparse convolutions, the O-Voxel kernels, CuMesh, nvdiffrast) each of which must match your exact Python, PyTorch, and CUDA combination. The custom node authors did heroic work shipping prebuilt wheels per configuration, but every PyTorch or CUDA update meant a new round of compilation failures, and installs regularly broke. This is now solved with the native integration in ComfyUI. Follow our installation tutorials for Trellis.2 and Pixal3D.

Licensing

Trellis.2’s own code and weights are MIT-licensed, but its original pipeline depends on NVIDIA’s nvdiffrast (for mesh rasterization) and nvdiffrec (for Physically Based Rendering), both distributed under the NVIDIA Source Code License which restricts usage to non-commercial research and evaluation. In practice, a studio couldn’t ship assets from the reference pipeline without stepping into a legal gray zone. These dependencies have been removed from with the native integration.

Then came Pixal3D

In April 2026, Pixal3D from researchers at Tsinghua University and Tencent ARC Lab got accepted at SIGGRAPH 2026. It pushed open-source 3D generation another step forward with its pixel-aligned generation establishing direct pixel-to-3D correspondences. The result is near-reconstruction-level fidelity to the input view, with detailed geometry and the same PBR material set.

Pixal3D is heavily built on Trellis.2 as it uses its backbone and shares its VAEs and DINOv3 image conditioning. This is why integrating it together with Trellis.2 made sense. However Pixal3D generally performs better than Trellis.2 as the generated 3D mesh strictly aligns with the input image.

Model highlights

Trellis.2

  • Single image to 3D asset. A 4-billion-parameter model that generates high-fidelity geometry and materials from one input image.
  • O-Voxel structured latents. A native, compact omni-voxel representation encoding both geometry and appearance, generating assets at effective resolutions up to 1536³.
  • Any topology. Handles open surfaces, non-manifold geometry, and fully-enclosed volumes.
  • PBR materials built in. A dedicated texturing model generates base color, roughness, and metallic maps.

Pixal3D

  • Pixel-aligned generation. Geometry is generated in direct correspondence with the input view. What you see in the image is what you get in 3D!
  • Explicit image back-projection. Multi-scale image features are lifted into a 3D feature volume, delivering near-reconstruction-level fidelity.
  • Cascaded refinement. A staged process progressively refines sparse structure, shape, and texture up to high resolution.
  • Built on Trellis.2. Shares the Trellis.2 backbone, VAEs, and DINOv3 conditioning.

What ships in this integration

The goal was simple: make the best open 3D models run in ComfyUI the way every image or video generation model does. A major thank-you goes to Kijai for the implementation, and to yousef-rafat for the initial draft this work built on. In addition to the native implementation, this has been an opportunity to make 3D generation a first-class citizen in ComfyUI. Here is what shipped:

Pure native implementation

Both Trellis.2 and Pixal3D now run as core ComfyUI nodes. The 3D post-processing that required compiled extensions has been reimplemented from scratch in PyTorch and SciPy. No nvdiffrast, no nvdiffrec, no per-configuration wheels, no PyTorch downgrade. If your ComfyUI runs, these models run on your current PyTorch.

Rebuilt 3D nodes

While these were shipped in an earlier version of ComfyUI, the Load 3D, Preview 3D, and Save 3D nodes have been rebuilt from the ground up to support these models and modern mesh workflows. We’re grateful to Terry Jia for his remarkable work on these nodes. Check out the nodes:

  • Load 3D (Advanced)
  • Preview 3D (Advanced)
  • Save 3D (Advanced)

Native mesh post-processing

Raw generative meshes are rarely production-ready, so this release introduces a new set of post-processing nodes:

  • Remesh Mesh: fixes holes and mesh imperfections.
  • Decimate Mesh: reduces face and vertex count to a target budget.
  • Smooth Mesh Normals: smooths the mesh volume.
  • Fill Holes: fill-in holes resulting from the generation
  • And more: Merge Meshes, Paint Mesh, Render Mesh…

A complete PBR texture set

Trellis.2’s texturing model generates base color, roughness, and metallic maps. Our implementation goes further: a new UV unwrapping node prepares the mesh for texturing, and two additional maps are generated: a normal map and an ambient occlusion map, both baked from the high-poly mesh. Are these textures perfect? No. But they’re free, generated on consumer hardware, and yours to use as you wish.

An honest word on quality

Let’s be direct: the best closed-source 3D generators (Hunyuan 3D, Tripo, Rodin) still produce better results than Trellis.2 and Pixal3D. If you need the highest quality and an API fits your pipeline, those remain strong options (all of them are available through ComfyUI’s partner nodes).

What this integration offers is different: the best open 3D generation available, running locally, at zero cost per asset, with no licensing restrictions on what you make. For iteration, prototyping, stylized work, 3D-to-2D workflows, and anyone who wants full control of their pipeline without spending an afternoon to install.

Getting started

  1. Update ComfyUI to the latest version 0.34.0 (or greater) or go to Comfy Cloud
  2. Download the workflows below, or find them in the template library.
  3. Follow the note in the workflow to download the models and save them in the correct model directory.
  4. Drop in an image and run.

Download Workflow

Model weights:


r/comfyui 14h ago

Workflow Included Use H3 To Replace Characters

Enable HLS to view with audio, or disable this notification

9 Upvotes

r/comfyui 18h ago

Show and Tell Testing My MiniMax-H3 → LTX 2.5 Upscaling Workflow — Results Are Looking Really Good

Enable HLS to view with audio, or disable this notification

17 Upvotes

I've been testing my MiniMax-H3 → LTX 2.5 upscaling workflow, and the results have been really promising so far.

One thing I've noticed is that the better your original MiniMax-H3 generation is, the better the final upscale will be. I'm getting good results even at lower resolutions, but faces still need stronger and more consistent input generations from MiniMax-H3 to maintain character consistency.

On my RTX 3060 12GB, the current upscale times are roughly:

  • 0.6 resolution: ~15 minutes
  • 0.8–1.0 resolution: ~20–30 minutes

It definitely takes some time, but I'm finding the results are worth it.

And of course, if you have a newer, more powerful GPU, you should be able to get even better results in less time, especially when pushing higher resolutions.

I was planning to release the workflow soon, but I want to spend a little more time testing it and seeing how much further I can improve it before sharing it.

So far, though, I'm really happy with how it's looking. 🔥

Would love to hear what you guys think and whether anyone else has been experimenting with MiniMax-H3 + LTX 2.5 upscaling.


r/comfyui 13h ago

Resource Introducing OpenGPEX — open-source browser image editor with ComfyUI integration. Generate, edit, export without switching apps

Thumbnail
gallery
5 Upvotes
I built OpenGPEX, an open-source image editor that runs in the browser. It connects directly to your own ComfyUI instance so you can run any workflow from inside the editor.

How it works:
1. Import your ComfyUI workflow JSON (or pull from server history)
2. Choose which parameters to expose (prompt, seed, model, etc.)
3. Hit Generate — the result comes back as a new layer
4. Edit it right there: background removal, adjustments, masks, text
5. Export as PNG/WebP/TIFF with transparency

It also reads embedded workflow JSON from any ComfyUI-generated image — you can view the full generation config and download the workflow file.

I wrote a step-by-step walkthrough using a D&D character portrait as an example: https://gpex.cloud/blog/comfyui-dnd-character-portrait

Open source. No signup needed. Fully self-hostable if you prefer to run it locally.

Try it: https://gpex.cloud
Source: https://github.com/gpex-cloud/opengpex

Let me know if you have any questions.

r/comfyui 5h ago

Workflow Included VIBRATE THE IMPOSSIBLE

Enable HLS to view with audio, or disable this notification

0 Upvotes

100% generated with MiniMax H3 using reference images and audio created with MiniMax Music 3 in ComfyUI Desktop, for the 2026 ComfyUI H3 Sync Sound Challenge.

Best of luck to everyone! Greetings from Argentina, and a huge thank you to MiniMax and ComfyUI for giving us this amazing opportunity!

Juan Manuel 🇦🇷


r/comfyui 23h ago

Resource Finally found a practical way to manage multi-shot MiniMax H3 videos in ComfyUI

31 Upvotes

I make short AI comic/drama videos and recently tried this open-source MiniMax H3 workspace:

https://github.com/siyuan-liu31/minimax-h3-video-studio

What surprised me is that it is not just another UI for generating individual clips. It actually helps manage a multi-shot video project.

I could:

  • generate each scene as a separate segment 
  • continue the next scene from the previous final frame 
  • use the previous video as a motion or character reference 
  • rerun only the segment that looked wrong 
  • keep prompts, references, settings, and results in one workspace 
  • merge the finished segments into a longer video

It also supports T2V, I2V, first/last-frame generation, multimodal references, V2V, and reference-assisted V2V.

The persistent workspace was probably my favorite part. Refreshing the page or disconnecting from a remote GPU does not mean rebuilding the whole workflow.

It is not a one-click hosted generator. You still need ComfyUI, the models, FFmpeg, and your own GPU. Character continuity between segments is not perfect either.

Still, for creating an actual story instead of a collection of unrelated clips, I found it surprisingly practical.


r/comfyui 5h ago

Help Needed Planning on upgrading my GPU, but, is going AMD worth it?

1 Upvotes

Hello!

I have an OK PC setup that I use for gaming + a bit of self-hosted AI stuff like personal chatbot. But it is struggling with GenAI particularly with video generation. I did some research before posting, and learned that AMD's ROCm has progressed enough that it's usable, but mostly when on Linux OS.

Before I bite the bullet and decide, I'd like to know more about it if anyone has tried or is using AMD GPU with video gen on ComfyUI.

My current PC specs for reference:

- 7800x3D

- 32 GB RAM

- 4060TI 8GB

Right now, it takes me at least 30+ minutes for just 2-sec WAN 2.2 video, not exactly sure what version. And 5-sec video is not doable.

My current options are:

- 5060TI 16GB - cheaper

- 9060xt 16GB - cheapest

- 9070xt 16GB - Ok price

I could probably try to squeeze out my savings and get a 5070TI 16GB, but probably not doable anyway, mostly due to prices and physical constraints on my PC. I just bought my current PC case this year, and I don't want to buy a new one just for my GPU.

I mostly use my PC for gaming, and use AI stuff not that often, but I do want my PC to be able to do AI stuff decently when I want to, at least do 5-sec videos comfortably, so I can learn more about how stuff works and all way better.

My main question is, is going AMD worth it now with the updates to ROCm? Or it would still be better to go with 5060TI, even if it's not much of an upgrade, just to get 16GB VRAM and proper support for AI stuff?

I also don't mind setting up dual-boot for Linux if that allows to use AMD GPU for AI better.

Thanks!


r/comfyui 6h ago

Help Needed How do I make ComfyUI use only the secondary SSD?

0 Upvotes

I use a portable version of ComfyUI installed on a secondary SSD. However, when running workflows, it uses storage on my primary SSD (C:) to create temporary files. I’ve been having issues because my primary SSD is full—I spent the whole day trying to figure out the cause of the errors, only to realize the drive was maxed out. Plus, constantly creating temporary files is going to wear out my primary SSD. So, I’d like to know how to configure it to use only the secondary SSD.


r/comfyui 14h ago

Help Needed Help with Inpainting

Post image
4 Upvotes

I am working on a project for school and not getting the results I hoped for. I am doing LoRA training for SDXL for WW2 equipment and want to be able to use them to inpaint the equipment on a map image. But I can not figure out what I am missing when it comes to the scaling of the image I want to create. For instance, I am testing this setup to insert a tank, not WW2 era, near the road but it always creates a metal blob unless I make the masked area huge. When I do that it is obviously way out of scale. Is there some other nodes I should include or change to fix the peoblem? Thanks.


r/comfyui 6h ago

Help Needed Workflow search...

0 Upvotes

I'm looking for a way to replace a specific object in a generated photo with an object taken from an image downloaded from the internet. If anyone already has a ready-made solution, I'd appreciate it if you could share. Thanks.


r/comfyui 7h ago

Help Needed Z-Image Turbo: hair still looks "plastic" despite prompt tweaks — what am I missing?

0 Upvotes

I'm generating photorealistic portraits with Z-Image Turbo on Comfy Cloud for brand/e-commerce content. Face and skin are coming out pretty solid already (pore texture, natural eye highlights, even lighting), but the hair keeps feeling like a solid plastic mass instead of real hair.

What I've already tried in the prompt:

- Specifying texture ("thick wavy hair, visible individual strands")

- Asking for matte instead of shiny ("no artificial shine, matte hair")

- Backlight/rim light to separate strands

- Translucency at the tips ("light passing through thinner strands, subsurface scattering")

- Varying strand thickness, loose flyaways crossing the face

Still not quite there. Question for the community:

Is this a limitation of Z-Image Turbo specifically (being a distilled/turbo model), and do I need to move to a heavier model? Is there a specific hair/texture LoRA you'd recommend? Or is there a postprocessing/upscaling node that specifically helps with this?

Any workflow, LoRA, or specific trick you've used to nail this would be hugely appreciated. Thanks!


r/comfyui 2h ago

Workflow Included Analysis Paralysis

Thumbnail
youtube.com
0 Upvotes

SyncChallenge