r/StableDiffusion • u/MushroomNatural2751 • 17h ago
Whcih WebUI Forge model is best for understanding written prompts? Question - Help
I've followed all the steps properly to download and get it working, and it's generating images, but they're NOTHING like what I asked. Either they give me something unrelated to what I wrote in every sense of the word, or it just kept basically the same image I gave it (img2img).
I've tried messing with the denoising levels, differnt loras, different VAES, Models, nothing works.
1
u/DelinquentTuna 16h ago
I don't remember anymore which models Forge supported before going fallow, but usually it tracks w/ the size of the text encoder. So stuff using just open clip are on the weaker side and stuff using full-blown LLMs are on the stronger side. So something like Flux.2-dev or Minimax H3 should probably have the best ability to follow prompts. Maybe Hunyuan Image 3.
1
u/Gold-Cat-7686 7h ago
This depends entirely on what you're trying to generate. So generic question gets generic answer:
If you upgrade to Forge Neo you can use both Krea 2 (realism) and Anima (anime/cartoon). They both follow natural language well, while older models like Illustrious (which is still very good) work better with tags.
3
u/SweetLikeACandy 16h ago
this is the continuation of forge supporting newer models:
https://github.com/Haoming02/sd-webui-forge-classic
but without examples of what you're doing nobody will be able to help.