An implementation of: Mastering Text-to-Image Diffusion: Recaptioning, Planning, and Generating with Multimodal LLMs
You need GPT4-Azure or Gemini Pro to use it. Local LLMs support is still being worked on.
https://github.com/zydxt/sd-webui-rpg-diffusionmaster
An implementation of: Mastering Text-to-Image Diffusion: Recaptioning, Planning, and Generating with Multimodal LLMs
You need GPT4-Azure or Gemini Pro to use it. Local LLMs support is still being worked on.
1 Comments
DrakeRichards@lemmy.world · 3 pts · 2y
Neat! I’ve known that Regional Prompter is powerful, but it’s too much of a pain for me to bother using. Hopefully this makes it easier.
tagginator@utter.online · 0 pts · 2y