• barsoap@lemm.ee
        link
        fedilink
        English
        arrow-up
        1
        ·
        5 months ago

        Image generation models are generally more than capable of doing that they’re just not trained to do it.

        That is, just doing a bit of hand-holding and showing SDXL appropriately tagged images and you get quite sensible results. Under normal circumstances it just simply doesn’t get to associate any input tokens with the text in the pixels because people rarely if ever describe, verbatim, what’s written in an image. “Hooters” is an exception, hard to find a model on Civitai that can’t spell it.

        • driving_crooner@lemmy.eco.br
          link
          fedilink
          English
          arrow-up
          0
          ·
          5 months ago

          I like to post sometimes on the “guess the song” AI communities, but more often than not, the Bing image creator just plast the lyrics on the image making it useless for the game.

    • MonkeMischief@lemmy.today
      link
      fedilink
      English
      arrow-up
      0
      ·
      5 months ago

      Yeah these look exactly like things I’d see on billboards in Vegas when certain conventions are in town…

    • drolex@sopuli.xyz
      link
      fedilink
      English
      arrow-up
      0
      ·
      5 months ago

      I will probably use these images in a corporate PowerPoint. I’m not asking for your permission, I’m warning you. Sorry, it’s too good. (I will credit you as a CTO of some company ending in -SYS or - LEA if you want)