← Back to context

Comment by Sharlin

3 years ago

Yes, currently SDXL doesn't really beat the best SD1.5 checkpoints quality-wise. But it (and the currently available checkpoints) shows awesome promise, so give it a six months or so.

The best 1.5 checkpoints are constrained in their output flexibility to achieve the quality they get though, and they don't follow prompts nearly as well as SDXL, so if the model doesn't naturally gravitate towards doing what you want it's very hard to steer it anywhere. SDXL also does a better job with full anatomy, which is the reason shared 1.5 generations tend to be torso up or portrait shots.

Currently SDXL is better than SD1.5 checkpoints at pretty much everything other than portraits (or anime drawings) of pretty women.

Unfortunately it seems that's all people want to generate, as is evident when you search for SD on Twitter.

  • Yes, point conceded, I should've said something about the flexibility and capability of SDXL rather than just image quality in a narrow sense.

  • Stable diffusion doesn’t grant the user a good imagination or taste unfortunately