Guy Parsons, author of the DALL-E 2 Prompt Book, discusses how text-to-image AI platforms require precise visual vocabulary rather than traditional graphical software controls.
Opinion
Parsons: Midjourney iterates on image models faster than OpenAI
“But then, of course, you have now tools like Midjourney, who've been, like, iterating on their text-to-image model, like, a lot more aggressively than OpenAI, who Understandably, I think maybe have some other things in the cooker, you know, which have now grow…”
Prediction Not checkable as stated
Parsons: Prompt engineering will be a specialized craft, not universal skill
“I don't think it will become, like, this necessary skill that everyone needs to have, but I do think it will become, you know, like, some people are expert wood whittlers or, you know, really good at Animating hair or whatever, you know, the people that develo…”
Insight
Parsons: AI image models struggle with precise spatial layout prompts
“They often describe generally what the image is about, but not like how you would draw it step by step. And that's why these tools are less good at saying like, I want this thing over here and then that thing next to it and then something on top and that thing…”
Insight
Parsons: AI prompts yield diminishing returns as length increases
“And I think there's something to be said, like, I think the longer they are, there's definitely, like, diminishing returns.”
Prediction Not checkable as stated
Parsons: Image-to-image AI models will power next-gen consumer interaction
“That's a really interesting space that's gonna probably power like the next generation of how people, especially consumers like interact with these products.”
Insight
Parsons: Generative AI fails when forced to produce highly specific work
“That's kind of the limitation of weather technology. Is at the moment, which is, it's amazing until you're trying to do something very specific. And especially if you want to do something very specific, this also to like a very high, like a professional standa…”