One thing I have to ask is why the AI insists on having slogans and lists on every object in its “illustrations.” You can see this everywhere. Hell, I went to a close friend’s funeral and they had AI generated memorial stickers. Even there, there are signs filling the negative space saying shit like “dedication, service, a safer community.” Is this an AI proclivity or do normal people just like that crap?
It’s essentially by design. To keep AI from destabilizing, the way AI is tuned is to essentially ‘go the same safe way’ on similar prompts. LLM models are essentially a giant field of weights for chains of tokens. Prompts are turned into tokens and the closest result is the output. The more room you give it to vary, the more different styles and variation it can express with a propmpt, but it raises the chance of destabilizing the ouput, producing things like additional fingers or complete garbage. It’s tuned to walk the same way to produce something we perceive as “passable” but it comes at the cost of taking away variation, which is why AI art/writing will consistenly converge towards the same narrow style set.
(This explanation is not super accurate. I just tried to convey the jist of it)
It’s hilarious what it does with a room full of table and chairs like a classroom. Cannot figure the legs, what a chair looks like, how people sit at desks, none of it.
This shit can’t even keep a consistent background over 4 comic panels
The scientist’s desk design changes in all four panels as well.
One thing I have to ask is why the AI insists on having slogans and lists on every object in its “illustrations.” You can see this everywhere. Hell, I went to a close friend’s funeral and they had AI generated memorial stickers. Even there, there are signs filling the negative space saying shit like “dedication, service, a safer community.” Is this an AI proclivity or do normal people just like that crap?
It’s essentially by design. To keep AI from destabilizing, the way AI is tuned is to essentially ‘go the same safe way’ on similar prompts. LLM models are essentially a giant field of weights for chains of tokens. Prompts are turned into tokens and the closest result is the output. The more room you give it to vary, the more different styles and variation it can express with a propmpt, but it raises the chance of destabilizing the ouput, producing things like additional fingers or complete garbage. It’s tuned to walk the same way to produce something we perceive as “passable” but it comes at the cost of taking away variation, which is why AI art/writing will consistenly converge towards the same narrow style set.
(This explanation is not super accurate. I just tried to convey the jist of it)
Still with the missing fingers too
It’s hilarious what it does with a room full of table and chairs like a classroom. Cannot figure the legs, what a chair looks like, how people sit at desks, none of it.