An experiment in which each TE is trained with different captions in FT.
1.L:characters tag
G:general tags
2.L:characters tag
G:characters tag+general tags
3.normal training
The method of giving one TE the perfect captions and the other only meta tags may be effective.
@Hosiokaa Yes, images from Nijijourney are part of the dataset. After a large-scale training, you might need to prepare a smaller dataset. This model was trained exclusively on Unet.
Merge attempts with WDXL didn't pan out. Maybe extensive training isn't needed, or perhaps using WDXL as a training foundation could yield better results.
#SDXL