r/StableDiffusion 37m ago

Discussion Testing DLSS 5

Enable HLS to view with audio, or disable this notification

Upvotes

Testing DLSS 5... Like many others, I was a bit confused about DLSS 5. I kept feeding it my hyper-detailed renders and only getting a color shift in return. After plenty of trial and error, I finally realized my mistake: this technology is developed to enhance video game graphics, so testing it on hyper-detailed renders makes no sense.

So, I generated a render in a 2020 video game style and started tweaking settings to find a final look with maximum effect, without worrying about flickering.

Final conclusion: What we have right now isn't very useful for us. Those of us using Latent Upscaler might be able to use it for color grading to get less saturated colors, but little else. Maybe in the future we'll get a DLSS 5 targeted at enhancing hyper-detailed graphics, but that's not the case for now.

Bottom line: If I want to generate a realistic render, I'll just generate it, there's no need to run it through DLSS 5.


r/StableDiffusion 1h ago

Question - Help How to better retain animation style for Ref2v?

Post image
Upvotes

I added video and image, which both are the same sources. I used an extension that automatically formats my text prompt., including copy over the animation style. I use H3 Prompt writer.

I am using the preset Workflow, but I replaced the text encoder with Qwen as an alternate due to memory issue. And I used I2v diffusion model instead of ref2v due to quality.

(No, I cant share the video example because it is not appropriate)


r/StableDiffusion 1h ago

Workflow Included It's a new look for the 90's!

Enable HLS to view with audio, or disable this notification

Upvotes

Made with Minimax H3


r/StableDiffusion 2h ago

Animation - Video Made A Professional short Animation video using Minimax-h3 (read description)

Thumbnail
youtube.com
3 Upvotes

Hey guys!

Since quite a few of you liked my previous videos, I decided to start a channel where soon I’ll be sharing tutorials and some of the workflows/tricks I’ve been using.

If you’re interested in learning how I’m making these videos, feel free to subscribe. I’ll be sharing a lot of the stuff I’ve figured out along the way, including:

  • My own workflows — free to download, with the tricks and settings I use
  • Character generation — how I use a Krea 2 character-sheet LoRA that I made to keep characters consistent, and how to get the style you want
  • Environment generation — how I generate environment images and then build scenes from them
  • MiniMax optimization — settings and techniques to make MiniMax faster while preserving quality
  • Video/audio tricks — ways to fix audio issues and continue a scene from the last frame to create longer sequences
  • Consistent voices — I also built my own UI app using BreezeTTS2 for voice cloning and generating consistent voices across an entire story

Everything I’m sharing is based on what I’ve been experimenting with myself, so hopefully it can save some of you a lot of trial and error.

If that sounds useful to you, you’re welcome to check it out!


r/StableDiffusion 2h ago

Discussion What would you consider to be the most consistent model at producing “consistent” images, non-realistic or realistic?

2 Upvotes

Could be actions, like “guy walking into store”

The same scene at different times of the day.

The same character doing different things.

You get the idea.


r/StableDiffusion 2h ago

Question - Help My First AI PC

0 Upvotes

Hi everyone, I have a question: I'm thinking of buying my first PC solely for AI. What minimum components do you recommend for running Stable Diffusion with Illustrious models? I've been using Free Google Colab to create images in Automatic1111 so I was thinking of buying a PC with similar specifications. What do you recommend?


r/StableDiffusion 3h ago

Discussion The underrated alternative to krea2 and ideogram4,guess the model?

Thumbnail
gallery
0 Upvotes

I like ideogram4 and krea2 a lot ,and I also really like this ONE.

What I personally prefer about it is the sense of depth and vastness that I don't feel as strongly in the other two.

Some notes from my testing:

Ideogram 4 can sometimes make skin and details overly sharp in a way that’s hard to fix naturally in post.

Krea 2 occasionally feels a bit static, like the subject was placed into the background rather than existing in the same space (it also uses Wan VAE which is the main reason I have also tested using the fp32 and realvae of it which solved texture to some extent).

These things can of course be improved with better prompting and Loras l,but tendencies are still there.

Just wanted to share a showcase of this model's capabilities.All in all i enjoy all the current models more the merrier.

They all deserve time,testing and appreciation.

some images are made using Boogu base for more creativity,some images are made using boogu Turbo for way greater prompt adherence


r/StableDiffusion 3h ago

Workflow Included Letting image-to-video artifacts compound into an impossible world

Enable HLS to view with audio, or disable this notification

8 Upvotes

Tools used: Gemma4 12b, LTX-2.3, Wan2GP, vibe coded video editor.

I’ve been experimenting with a slightly self-destructive image-to-video workflow where continuity comes from letting the model reinterpret its own mistakes.

I started with an almost completely black image with a few faint stars, then gave Gemma4 12B the track’s beat grid and energy-shift analysis, along with a long description of the overall concept: a monolith, a hallway of impossible geometry, and a progression from restrained movement into increasingly unstable architecture.

Gemma4 wrote all 27 scene prompts beforehand.

For generation I used LTX 2.3 with the audio-reactive LoRA. I also tested LTX 2.5, but for this workflow it became too artifact-heavy too quickly. LTX 2.3 held the scene structure together longer while still producing enough weirdness to evolve in interesting ways.

The process was simple: generate a clip with the correct audio slice, cut it on the beat grid, then take the frame immediately after the cut and use that as the starting image for the next generation.

The fun part was deliberately keeping some “bad” transition frames.

If a flash landed on the frame used for the next clip, the model might reinterpret it as a permanent light source. A lens flare could become a horizon or an entire landscape. A warped piece of geometry that only existed for one frame could become a major architectural feature in the next scene.

So the artifacts compound.

Eventually the video loses any reliable sense of scale or orientation. Surfaces become spaces, structures fold into other structures, and at some points I wanted an Inception-like feeling where you can’t tell which way is up, or whether the camera is traveling deeper into the structure or pulling outward into something much larger.

The audio-reactive LoRA helps hold it all together. Even when the geometry becomes increasingly strange, the environment keeps breathing, unfolding, compressing and reorganizing itself with the growing low end.

What I like most is that the continuity doesn’t really come from visual consistency. It comes from causality.

Every scene inherits some accidental information from the previous one, and the next generation has to decide what that information actually is.

After enough generations, the model is basically building a world out of its own misunderstandings.


r/StableDiffusion 3h ago

Animation - Video Hope my humble work would inspire the low vram folks!

Thumbnail
youtube.com
7 Upvotes

An AI-assisted webcomic creator here. I'm among the vram and ram-poor folks, with my humble RTX 3060 12 GB vram and a mere 16 GB ram. Since the beginning of time, I've convinced myself that comic is my focus, and so what I have is enough. I don't want to pay any opportunistic video gen platforms out there. Don't want to rent GPU and trouble myself with transferring assets and models from storage to storage. Aside from light experimentation, I had thought I'd stay away from video gen for a very long while.

That is, until the arrival of Minimax H3... And just two weeks after setting it up (ComfyUI, default ref2va and fl2va workflows), I was able to edit together an animated trailer for my webcomic on my own machine, *entirely local*! Granted, in terms of generation quality there's a lot to be desired, as any resolution beyond 0.4 mp is too slow for me to comfortably iterate on. But still, oh such *feeling* when the world I built suddenly came alive for the first time, and on my own machine, too!

Feel free to ask me anything. Happy to share.


r/StableDiffusion 3h ago

Workflow Included Super nothing!

Enable HLS to view with audio, or disable this notification

12 Upvotes

Made with Minimax H3


r/StableDiffusion 4h ago

News MiniMax H3 acceleration arena/leaderbord: 15+ H3 LoRAs, fine-tunes, Max

Thumbnail
huggingface.co
190 Upvotes

Hey folks, I've built an so we can have a proper leaderboard on 15+ different LoRAs, fine-tunes and acceleration technique. Baseline is included for anchoring, and M3 Max is also included given the promise to open source

There are there being compared: H3 baseline, FastH3 family, H3 Acc family, Lightx2v family, Larryvrh family, JoyFox family, RAVEN, FlashGen, TuTu, SilverOxides merges, Plaguekind merges and Fal's H3 Max


r/StableDiffusion 5h ago

Discussion Thoughts & opinions on Anima - turbo-v1.1

Thumbnail civitai.red
6 Upvotes

So I've been using the new Anima turbo-v1.1 model and I have to say it's pretty good now and then. The thing I like is its unpolished look like it does have a rough default art style in my opinion, but I kind of like that as it looks less too polished. It also has pretty good diversity as well. It also works pretty well with LORAs like the base model, however I haven't tried multiple LORAs together.

What I don't like about it is it can be a little inconsistent regarding prompt adherence and also Anatomy and sometimes it can give it for the subject extra or missing limbs and miss out details/objects in the prompt sometimes. Not often but sometimes. To be fair, the turbo model also has the same issue as well sometimes and is probably due to the low CFG and low steps of turbo distilled version.

It looks like a bit of an improvement to the previous version, but I do hope the anima team works on a bigger and stronger turbo Lora for the base model as it's still much better especially when using other fine-tune anima Checkpoints plus better Lora support.

I'm curious to see what you guys think of it as it is a fairly new release.


r/StableDiffusion 5h ago

Resource - Update Infinite AI Twitch Streaming Project (2xB200 480p Minimax FastH3)

Thumbnail
youtube.com
0 Upvotes

far from a perfect setup, but this is a interesting project for sure. would not mind the b200 prices to go down tho. framework i used ish https://github.com/reactor-team/infinite-livestream. 2xb200 + gpt-5.6luna


r/StableDiffusion 5h ago

Question - Help Looking for easy free way to run comfy ui at the cloud ?

3 Upvotes

My laptop doesn't powerful enough to run minimax H3 local so i need easy way to run minimax h3 on the cloud . I already try few methods like Google collab but fosent work and always keep making sever error and the other comfy ui clouds in different site dosen't load very properly. So yeah if there's any ways to run comfyui online for free or minimax h3 local i will appreciate it


r/StableDiffusion 5h ago

Question - Help One of these nodes, Spectrum or Sage attention, effed up my 5070ti so bad that the system sad I have no GPU even after a cold restart

Post image
0 Upvotes

I am quite sure one of these two nodes is the culprit. Most other workflows seem to work fine but this workflow led to the gpu completely being knocked out of my system. A restart got it back but after the 10th time or so even the restart would not bring it back. Took another restart. Now when I start the workflow without bypassing these nodes, the fan hits the ceiling from 0 to 100% in 2 seconds and I get fully black screens with the fan running ad infinitum.

I tested GPU and VRAM for 10 minutes each with OCCT per Claude recommendation and there are no errors. Checked the event manager protocol but it does not show anything relevant for the last two days when I had the issue. My GPU hardly ever goes over 75° C and it didnt seem to when I crashed too.

I deinstalled the nvidia drivers with DDU and installed the newest one (which Claude told me afterwards is apparently not stable).

I took out the GPU, checked connections, had a more knowledgeable friend look at the hardware. There seem to be no issues. What could Sage or Spectrum even do to cause something like that?

It did not make a difference if I started in --low vram or not, both times the gpu crashed.

Nvidia Version: 616.56 (the one I had before caused the same issue though)

Cuda 13.4 (Cuda 13.3 or what I had before caused the same issue)

ComfyUI 0.34.0

Edit: Happened without Sage attention now too .. it is a different issue. Will try throttling GPU now. Officially it is only 75°C but HWinfo estimates Hot Spot at over 107°C


r/StableDiffusion 6h ago

Question - Help Which image generator is used for these images?

Thumbnail
gallery
0 Upvotes

Any idea? Is this midjourney?


r/StableDiffusion 6h ago

Question - Help High fidelity Videos using H3

4 Upvotes

What methodology or workflow do you follow in ComfyUI to achieve high-fidelity video?

When I say high fidelity, I don't just mean preserving faces—I mean maintaining fine details across the entire frame, including objects, textures, clothing, architecture, and background elements.

I'm trying to get something closer to Seedance 2.5 level quality using H3, if that's realistically possible. I've been doing a lot of trial and error, but so far I'm only getting somewhat good results by increasing the MP/resolution. Even then, the output still feels like it's interpolating or hallucinating low-quality details rather than actually generating high-fidelity detail.

Is there a specific workflow, model combination, sampling strategy, or refinement/upscaling pipeline you recommend for this? My main obstacle right now is that low-quality/interpolated details keep appearing throughout the video, especially in the background and secondary objects.


r/StableDiffusion 6h ago

Question - Help Prompt or clothing problem

1 Upvotes

What do i do wrong? When i have a female character and she wears like a shirt her breast shrink.

I work in Krea 2.

It looks like the clothing preventing the anatomy. Or, i don't really know.

Thanks


r/StableDiffusion 6h ago

Animation - Video I didn't know bigfoot visited my kitchen and grabbed the tomatoes (H3)

Enable HLS to view with audio, or disable this notification

4 Upvotes

So it's the first time I finally am fully satisfied with the quality of my H3 videos and it's thanks to the new 3D Latent Upscaler of H3 of HuggingFace.


r/StableDiffusion 7h ago

Animation - Video Inuyasha Love Triangle Solved

Enable HLS to view with audio, or disable this notification

12 Upvotes

A silly idea I had that I hope you guys had a good laugh at. Still love this classic anime!


r/StableDiffusion 7h ago

Question - Help How to improve quality when using H3 with Turbo? (RTX 3000 6GB VRAM)

0 Upvotes

Video link: https://streamable.com/irnjvm

Check video above.

Running MiniMax H3 Turbo on a spare laptop with Quadro RTX 3000 6GB VRAM and 64GB RAM.

352×608, 8 steps, Euler + Beta, Turbo LoRA @ 1.0. Takes ~550 sec for a 5-sec video.

I know the GPU is very limited 😅 Quality isn’t great, but it runs without OOM, so I’m wondering if I can push it further.

Any tips on what to do for better quality? Except for buying a new GPU.


r/StableDiffusion 7h ago

Workflow Included Personification:Planets (tarot cards)

Thumbnail
gallery
25 Upvotes

r/StableDiffusion 7h ago

Question - Help Looking for Grok img2img Alternative in Local

Thumbnail
gallery
14 Upvotes

Are there any Local Models that can achieve this level of Natural-ness and Realism, not over texturing and over crisp images ? I've been looking for a while and can't find any Closer to this, these images used Grok img2img for Lighting, skin Texture and Overall phone Shot vibes, the base images generated by Local SDXL/Illustrious For the Semi Realistic look, and i used Grok (The Last Grok model before the update), to improve realism, pure img2img and not even a slightest angle change made by Grok, since the Last Grok update everything turned to crap, Everything looks worse and So AI Plastic


r/StableDiffusion 8h ago

Question - Help Minimax H3 Ref key words help

10 Upvotes

I've read through the prompt guide, but I'm still having some trouble understanding when to use which of these

fully_preserved, partially_preserved, attribute_transfer, weak_reference

From what I understand you use these in the retention_analysis block. Let's say I want to fully_preserve the face, hair, and body characteristics from <Picture 1>, but I want to swap the character to wear the clothing from <Picture 2>.

Do I use

<Subject 1> (appears in [Shot 1], [Shot 3]): fully_preserved - and describe the portions of the picture I want to fully_preserve? 
<Picture 2> ([Shot 1] first frame): fully_preserved - and describe the clothing I want to fully preserve?

or do I 

<Subject 1> (appears in [Shot 1], [Shot 3]): partially_preserved - because I want to change the clothing she's wearing?
<Picture 2> ([Shot 1] first frame): partially_preserved? Or attribute transfer?

r/StableDiffusion 8h ago

Resource - Update Just tried ChaiNNer for the first time. It's a node based upscaler app with many other image processing uses. I installed it to test out a new map upscaler that looked interesting. I really like its node menu layout on the left of the GUI. Thought I'd share in case anyone is interested.

2 Upvotes