REVIEWS / AI IMAGE / OWNER INSIGHTS

🦉 WE READ 226 OWNER COMMENTS

DALL-E 3: what owners actually say

Owners praise DALL-E 3's prompt-following and text rendering but note visual quality lags behind Midjourney and SDXL

LEMMY · 90 HACKERNEWS · 75 YOUTUBE · 29 PRODUCTHUNT · 18 REDDIT · 10 STACKEXCHANGE · 4

What owners complain about

  • Visual quality below competitors COMMON

    Multiple users note that while DALL-E 3 follows prompts better than MJ or SD, the actual output quality is not as visually impressive — one commenter summarized it as 'better at following prompts than MJ or SD. But the output quality not so much.'

  • Cherry-picked marketing vs. real results SOME

    Users carry skepticism from DALL-E 2, which they say used heavy cherry-picking in marketing while the actual product was 'fairly terrible and difficult to get any good results.' Several worry DALL-E 3 demos may follow the same pattern.

  • Falls apart with specific vision in mind SOME

    One tester observed that AI generators look great in cherry-picked showcases or with low expectations, but 'the more you have a specific image in mind or the more you want certain details the worse it gets. And you quickly realise the limitations.'

  • OpenAI silently modifies prompts SOME

    Users discovered OpenAI injects an internal prompt that forces diversity into depictions of people and removes requests for specific descent, which overrides user intent without disclosure.

  • Long prompt engineering is tedious SOME

    One commenter found generative art 'really frustrating' and felt typing long prompts was 'not worth the effort,' while another noted the 'create text version of image' intermediate prompt matters enormously and requires extensive tweaking.

What owners love

  • Best-in-class prompt adherence

    Users consistently note DALL-E puts the things you actually asked for in the image, unlike SD which 'sometimes no matter what you tell it to do, it just doesn't do it.' Even critics concede this advantage.

  • Text coherence in images

    Multiple commenters were blown away by text rendering, calling the avocado example 'insane' and saying they had 'yet to see any image prompt AI get text and composition anywhere near this level.'

  • Accessible to non-artists

    A software engineer with 'absolutely no eye for design' reports being able to write a prompt and get a 'good-enough image in 90% of cases' that was previously unreachable without hiring an artist.

  • ChatGPT integration lowers barrier

    Users appreciate that ChatGPT translates natural language into effective image prompts, making it significantly easier to use than crafting raw prompts for SD or MJ.

Surprising patterns

  • The intermediate step matters more than the user prompt: users found that GPT-4V's internal 'create text version of image' prompt heavily shapes output, and tweaking that hidden layer produces dramatically different results from the same input.
  • Prompt position swapping affects output: one user ran 10-iteration tests swapping subject order in prompts and found the position of each subject in the prompt changes what the model emphasizes.
  • Professional artists in the thread don't see DALL-E 3 as their tool of choice — they point to ControlNet as superior because it lets you 'exactly determine the visual structure,' suggesting DALL-E 3 is optimized for non-artists rather than working creatives.

WHO SHOULD SKIP IT

Professional visual artists who need precise structural control over compositions, or anyone with a highly specific image already in mind — multiple owners report results degrade the more detailed your expectations become.

5.8/10 GYIBB verdict
Full review →

Synthesised from 226 real owner comments across 6 platforms. Every point is grounded in the comments — no marketing, no AI guessing. How we do it →