this post was submitted on 08 May 2024
1131 points (98.7% liked)

Science Memes

10348 readers
1477 users here now

Welcome to c/science_memes @ Mander.xyz!

A place for majestic STEMLORD peacocking, as well as memes about the realities of working in a lab.



Rules

  1. Don't throw mud. Behave like an intellectual and remember the human.
  2. Keep it rooted (on topic).
  3. No spam.
  4. Infographics welcome, get schooled.


Research Committee

Other Mander Communities

Science and Research

Biology and Life Sciences

Physical Sciences

Humanities and Social Sciences

Practical and Applied Sciences

Memes

Miscellaneous

founded 2 years ago
MODERATORS
you are viewing a single comment's thread
view the rest of the comments
[–] [email protected] 8 points 4 months ago

Image generation models are generally more than capable of doing that they're just not trained to do it.

That is, just doing a bit of hand-holding and showing SDXL appropriately tagged images and you get quite sensible results. Under normal circumstances it just simply doesn't get to associate any input tokens with the text in the pixels because people rarely if ever describe, verbatim, what's written in an image. "Hooters" is an exception, hard to find a model on Civitai that can't spell it.