Comment

CluckN@lemmy.world ⁨1⁩ ⁨year⁩ ago

I’m surprised it’s able to make readable text.

Sort:hotnew top

zalgotext@sh.itjust.works ⁨1⁩ ⁨year⁩ ago

Blovw the competittio

source
- Aggravationstation@feddit.uk ⁨1⁩ ⁨year⁩ ago
  Blovw it hard!
  
  source
barsoap@lemm.ee ⁨1⁩ ⁨year⁩ ago
Image generation models are generally more than capable of doing that they’re just not trained to do it.

That is, just doing a bit of hand-holding and showing SDXL appropriately tagged images and you get quite sensible results. Under normal circumstances it just simply doesn’t get to associate any input tokens with the text in the pixels because people rarely if ever describe, verbatim, what’s written in an image. “Hooters” is an exception, hard to find a model on Civitai that can’t spell it.

source
ghen@sh.itjust.works ⁨1⁩ ⁨year⁩ ago
The worst part is that it now tries to add text to a whole lot of pictures

source
SomeAmateur@sh.itjust.works ⁨1⁩ ⁨year⁩ ago
Yeah it’s been improving over the past few months. It’s hit or miss though

source
- driving_crooner@lemmy.eco.br ⁨1⁩ ⁨year⁩ ago
  I like to post sometimes on the “guess the song” AI communities, but more often than not, the Bing image creator just plast the lyrics on the image making it useless for the game.
  
  source