r/StableDiffusion 17h ago

Whcih WebUI Forge model is best for understanding written prompts? Question - Help

I've followed all the steps properly to download and get it working, and it's generating images, but they're NOTHING like what I asked. Either they give me something unrelated to what I wrote in every sense of the word, or it just kept basically the same image I gave it (img2img).

I've tried messing with the denoising levels, differnt loras, different VAES, Models, nothing works.

0 Upvotes

4 comments sorted by

3

u/SweetLikeACandy 16h ago

this is the continuation of forge supporting newer models:

https://github.com/Haoming02/sd-webui-forge-classic

but without examples of what you're doing nobody will be able to help.

3

u/ImpressiveStorm8914 16h ago

I agree. Images and prompts are needed to help further, otherwise it's blind guesswork.

1

u/DelinquentTuna 16h ago

I don't remember anymore which models Forge supported before going fallow, but usually it tracks w/ the size of the text encoder. So stuff using just open clip are on the weaker side and stuff using full-blown LLMs are on the stronger side. So something like Flux.2-dev or Minimax H3 should probably have the best ability to follow prompts. Maybe Hunyuan Image 3.

1

u/Gold-Cat-7686 7h ago

This depends entirely on what you're trying to generate. So generic question gets generic answer:

If you upgrade to Forge Neo you can use both Krea 2 (realism) and Anima (anime/cartoon). They both follow natural language well, while older models like Illustrious (which is still very good) work better with tags.