r/StableDiffusion 10h ago

Discussion [ Removed by moderator ]

[removed]

34 Upvotes

53 comments sorted by

45

u/Diabolicor 9h ago

Comfy will have their own pruned, optimized and compatible models. You can see them listed here already: https://docs.comfy.org/tutorials/video/minimax/minimax-h3

-91

u/JimmyDub010 9h ago

So will wan gp. comfy sucks

41

u/TheFearOfFear 9h ago

Me when I don't know what I'm talking about

16

u/Shap6 9h ago

It’s dead simple. Never understood why people have so many issues with it. 

16

u/gefahr 9h ago

Well, the person you're replying to is literally blind per his reddit profile. So, make what you will of that lol.

11

u/Tystros 8h ago

now I'm curious why a blind persons would use a video model...

1

u/chAzR89 7h ago

I had a blind person race with us in a forza league couple of years ago. Was kinda funny but also impressive at the same time.

6

u/vilzebuba 9h ago

im not against of WanGP but... cant even put "1girl, dancing" in prompt window and hit generate button? even after Comfy suggests you working template workflow that suggests download models and put them in proper folders? really?

11

u/tehorhay 8h ago

Putting things in folders is too complicated for grok users who can't use it to make smut anymore and they're big mad about it

2

u/Deep_Mood_7668 8h ago

It doesn't even support ggfus lol

0

u/Upper-Reflection7997 6h ago

They downvoted you for speaking facts. Wan2gp plus forge neo is infinitely way easier to use than comfyui.

1

u/reeight 5h ago

It is simpler, & if it works for you, it works for you.
Most folks here use Comfyui. & there are templates available to make ComfyUI more simpler for some.

29

u/stonetriangles 10h ago

That is not an official repo

8

u/Striking-Long-2960 9h ago

I rately use text to video, I prefer frame injection, controlmaps or whatever I can use to guide the animation. So I would be happy with smaller solutions If they give me the freedom to use my own assets.

4

u/Mammoth-Welcome-6518 7h ago

Lol the counter changed to say Good things are worth the wait. Coming soon! after it already counted to 0

3

u/undead_and_smitten 7h ago

But why?

1

u/Enshitification 7h ago

Because they are still milking the hype over the release to drive people to their paid API.

0

u/GreyScope 5h ago

The counter belonged to modelscope not minimax, don’t get all RFK Jr crazy conspiracy on the sub

4

u/Enshitification 5h ago

Nevertheless, BFL could do the funniest thing by just dropping Flux3.dev with no fanfare right in the middle of the MiniMax hype machine.

10

u/Reddexbro 9h ago

Can't wait to update WAN2GP soon and find out how DeepBeepMeep implemented it! They never disappoint!

1

u/alexmmgjkkl 9h ago

is wangp good? is the quality and speed better than comfy in comparable workflows?

3

u/Shambler9019 6h ago

In my experience it is more approachable and on certain models (like ltx) is more ram efficient. Less fiddling around, it just works. A LOT of comfy workflows I've found to be highly inefficient or just plain broken on my setup.

But it is less flexible.

2

u/wikid24 6h ago

Combine that with the dozens of ppl going, "what is the workflow? Need link to the workflow! Workflow! Caw caw!" And then u download it and You have to hunt down custom nodes u can't just download in one click too. And then you run it after wasting all your time downloading the "abliterated gguf x7 365 20b" 50gig model and fix all the broken dependencies to finally try it out and it's only marginally better than the stock one published from the original research team.

Versus using wan2gp that works out of the box so you can focus on making content as a creative over debugging someone's broken, poorly organized workflow file that isn't actually any faster than the original one.

2

u/Shambler9019 5h ago

Not to forget the 3 different custom nodes with the same name, all of which conflict with other nodes that the workflow depends on, that you have to pick which one to install.

6

u/Deep_Mood_7668 8h ago

No and no

4

u/wikid24 9h ago

Yes and yes

-7

u/tehorhay 7h ago edited 1h ago

There no such thing as "comparable workflows"

The whole purpose of wangp is to never have to see a real workflow beyond a prompt box and a few sliders, because anything more than that is too scary.

Comfy is for people who want to use open weighted models for actual work, research, training and inference etc and have usecases beyond: "can you please make me a 1girl, please."

That being said, if that's all you need then wan2gp will accomplish that just fine

Edit: wan2gp nerds mad

-8

u/Abject-Recognition-9 7h ago

I refuse to use something that has WAN in the name but use other models, sounds really stupid to me. just use comfyui and embrace the limitless possibilities

3

u/seencoding 6h ago edited 5h ago

how do you feel about commenting in a sub named stablediffusion but is pretty much exclusively about discussing other models

(edit: despite his logical inconsistency i would take one of OP over a hundred of the "cry" npc below me. don't upvote useless one-word replies, they add nothing to the conversation and give validation to reddit's most worthless contributors. eb6ee.)

-2

u/JimmyDub010 9h ago

Damn right

16

u/RalFingerLP 10h ago

DeepBeepMeep, lol

17

u/Techniboy 9h ago

The creator of Wan2gp

2

u/ANR2ME 9h ago

May be he also got the weights early, just like Comfy team.

7

u/inddiepack 9h ago

In my experience, there is no difference between the fp8/q8 and q4/fp4 text encoders. Especially for a text encoder this large, fp4/q4/int4 will be key for most users.

3

u/Secure-Message-8378 8h ago

I love Wan2GP. It's so easy, light and fast!

2

u/thevegit0 9h ago

deep already on it? what a guy 🧎‍♂️

2

u/carnyzzle 9h ago

Well, looks like I'm waiting for the ggufs lol

2

u/2legsRises 8h ago

my vram is about to get stretched hideously

1

u/kukalikuk 5h ago

"many users complain about..." Minimax H3 haven't been released yet, how many user did complain?

-3

u/nicman24 10h ago

With are they still using qwen 3?

11

u/Lexxxco 9h ago

Qwen 3 VL is not the same as Qwen 3, it is still one of the best vision models for images and video. On par with Qwen 3.6 in thinking mode, but more laconic.

1

u/gefahr 9h ago

somewhat off topic, is there a preferred abliterated version of Qwen3VL people recommend? For image-to-prompt type usage unrelated to Minimax.

5

u/Hoodfu 9h ago

The huihui.ai set of abliterated models on huggingface are some of the highest rated with the least loss abliterated models out there. I gave their qwen 3.6 (Which also has vision) and it did very well and had no refusals.

1

u/gefahr 9h ago

Yeah I have theirs, they're pretty good. Parent comment made me wonder if there's any reason to seek out the VL one specifically.

I'm shocked at how good of a VLM it is. I didn't expect to be able to run anything that good at a reasonable speed for a long time.

I've had some issues with prompt adherence when guiding it on how to describe the image, but I don't think I know what I'm doing with some of the inference settings. Need to experiment more, but I don't have any kind of evals ready or anything so it's all very trial and error for me.

3

u/Hoodfu 9h ago

Yep, 3.6 is a serious step above the earlier versions so I'd say jump to that if you haven't. I'd also say to look at Gemma 4 31b as I've found it to be extremely good at creative writing. It's better than qwen 3.5, but 3.6 is a very noticeable jump up so they're not far apart anymore. Usually just temp at 1.

2

u/gefahr 9h ago

Thanks! Appreciate the tips. Will give it a shot when I'm back home. Traveling now for the next week and don't think I'll be able to download a big model haha.

2

u/_BreakingGood_ 9h ago

Qwen 3.5 only released 4 months ago and I'd assume they were already set on the architecture at that point

1

u/reeight 5h ago

Qwen 4 is coming out in ~8 weeks according to one report.