• lime!@feddit.nu
    link
    fedilink
    arrow-up
    4
    ·
    7 days ago

    i don’t think the distinction “image model” really holds anymore. most big language models can generate images on their own if you bully them enough. they’ve ingested enough binary data to know what a valid jpeg looks like and you can use that to get base64-encoded images out.

    as for banning diffusion models rather than transformers… well we have text diffusion models now.