AI Image Users Fighting Unpredictable Output
Users are frustrated with the unpredictable quality and lack of control in AI image generation tools like ChatGPT and Gemini. They struggle to achieve desired results, encountering issues like illogical outputs, artificial-looking images, and text errors, leading to wasted time and unmet expectations. The evolving landscape of models and tools adds to the confusion.
SOURCES (60)
“But that doesn’t mean they are the same, which is the point I was making. Just my because they share the basic function of opening a word document, they can still be completely different. Of course things appear the same if you don’t understand them…”
“I want to keep it simple. In my view this would confuse most users as it’s not gonna really work like that if you have the same seed and simply replace a red dress with yellow dress in the prompt it’s gonna be the same image. If there is enough demand, I am more thinking about adding a feature where you paint the regions which should stay un-changed.”
“Is there a way we can get the seed, so we can try and recreate the same image in different poses, outfits, etc?”
“okay that's hilarious thanks for sharing, i had to try it myself: https://chatgpt.com/share/6a5e6090-f454-83ed-8d1e-e17ad3526899 gpt 5.6 sol high, custom instructions so you can see nothing weird was done: "System Instruction: Directive Mode (Friendly, Slightly Warmer Version) Prioritize clarity, precision, and high information density. *always* use the highest possible reasoning effort to ensure as high as possible accuracy during any interaction at all Assume the user has strong cogni”
“Too many highlights and looks a bit cartoonish.. 🤔 not sure what style I was going for, but the result just looks like ugly AI art. Do I try to make it look more realistic? How would one do that? submitted by /u/Onextto0 [link] [comments]”
“This made me realize my post undersold what kind of assets we're talking about, so that's on me. It's not parametric icon generation where style = stroke width + radius. It's closer to illustration packs: characters, scenes, a visual genre. The core job is "here's one image, now give me ten more in the same hand" - same character in different scenes, same rendering style across the set. Which is why your six-assets test actually fits even better than you intended. F”
Image Generation broken in all our custom GPTs since the new default model — the GPT returns the internal /mnt/data/…png path instead of rendering
“Huh. That's weird. I just have the 20 dollar subscription and it works fine. I have it installed on my desktop rather than an extension. Might also be a region thing? I'm in the us”
“Does anyone have any luck generating images with a true transparent background? No matter what prompting, it just makes a fake checkerboard background instead of a true transparent one. I tried: "alpha transparency, true transparent PNG background, no baked-in checkerboard" ...and the likes, but no luck. At the same time people on reddit and youtubers all claim that this is possible. I tried their literal prompts, but no luck again. That's a big drawback to my use case for an other”
“https://preview.redd.it/1gr7xu20k0eh1.png?width=1722&format=png&auto=webp&s=30dfbdb1861276987777627b4eaddf86be8a407e https://preview.redd.it/rfp5owx1k0eh1.png?width=1668&format=png&auto=webp&s=2e9d7c41ec60024cec5e41487fe62487e9592146 just did a preview practice exam and do these two explanations make sense? I feel like they are basically the same scenario, but contradictory. I would love an explanation! submitted by /u/WordOk_ [link] [comments]”
“The worst part is that it uses up my image generations. submitted by /u/Disastrous_Lie_6698 [link] [comments]”
“Wait so you just developed this pose code syntax but the llms are already able to write in it?”
“https://preview.redd.it/ebpqpv7oqzdh1.jpg?width=1920&format=pjpg&auto=webp&s=d7f26e3d53ee8c9a85dcb5b0a80894423e65f34d submitted by /u/yahbluez [link] [comments]”
[ChatGPT] ★★★ 3/5 (v1.2026.188) — I uploaded two eye photos and asked for those exact photos to be edited. The image generator ignored my uploaded images and repeatedly created completely new eyes…
Custom GPT understands Knowledge but cannot use official templates and logos when generating images
“https://preview.redd.it/ffzzqqkp0wdh1.png?width=2000&format=png&auto=webp&s=78b1098acb67dc846771d9839b7ee679838ea9e8 album cover + point-and-click + abandoned garden = this submitted by /u/Able-Buy-4516 [link] [comments]”
“Any feedback at all would be appreciated! My brain is mush right now 😵💫 submitted by /u/TinTinV [link] [comments]”
July 2026 (Theme: Perspective) — ChatGPT / API Image Generative Art Gallery, Prompt Tips, and Help
“Execuse me if this turns out to be long,so apparently when I ask chatgpt plus to make an image,it successfully makes it the first try but when I click on the retry button it gets stuck generating it forever and after long minutes it says "stopped" everytime I have to open a new chat to fix this I tried deleting all chats and discussions from data controls and it still persists,I even cleared cache and data and it still happens Knowing that this only happens on the app mobile version of”
“Well yeah, because music is not a modality of the models involved at all. It's literally just an LLM with the timestamps of each line injected into its context, and then making an API request to image / video generators. I don't really understand what the point of this project even was; stitching together dumber models with bespoke glue logic like this takes us further away from general intelligence and confuses the public.”
“This is what i'm doing with the tiles, but it doesn't fell right, what can i do to make it seems really good? And like blasphemous1/2? submitted by /u/mori_01011 [link] [comments]”
“https://preview.redd.it/d19cocm3tldh1.png?width=1065&format=png&auto=webp&s=d1f8eb4eeccb0ab9718077cc5f97874cf4226576 Me too...”
“It helps with the vibe, imo. It really depends on the crop & how artistic it is. Don’t be afraid of using generative fill to manipulate dead space (if the contract allows of course). Also look at photos that didn’t make the cut the first time, as they also may have more potential when the model’s face is cut out. There is the risk that users think it’s ai generated, but if hands & details/background look normal, you should be fine IMO.”
“Specifically Qwen3.6-35B-A3B-uncensored-heretic-Q8_0.gguf, temp 0.0, "Create an SVG of a Darth Vader." (In general text use, anything 4 experts or less seems to seriously break down. (8 is the default for this model). submitted by /u/temperature_5 [link] [comments]”
[ChatGPT] ★ 1/5 (v1.2026.188) — OK but not good , creat Promts not good alway wrong n not correct, I have to correct n fix a more than 10x image still not…
Hey /u/CanYouEvenPhoto , If your post is a screenshot of a ChatGPT conversation, please reply to this message with the conversation link or prompt. If your post is a DALL-E 3 image…
“The actual flow: had GPT 5.6 Sol write out the detailed prompt in the browser first, something like set up Blender MCP, model a MacBook floating in empty space, soft downward reflection, key light from the upper left, then render it as a clean product shot, and pasted that straight into Cursor to run it. Same model the whole way through, one endpoint: https://www.atlascloud.ai/models/openai/gpt-5.6-sol”
“I wasn't really knocking it haha. From what I understand is, image generation isn't really it's main purpose like say running something in stable diffusion huh? So here, is it more like asking the model to handmake you something in ms paint?”
“SVG generation, not image generation, does not have the same image generation capabilities that ChatGPT has.”
I believe I may have fixed it. It was messing up and calculating essentially backwards on raw pasted images. Could you try again please and let me know the outcome?
“Combining multiple images from ChatGPT in Photoshop. https://preview.redd.it/3zxcayto1gdh1.jpg?width=7200&format=pjpg&auto=webp&s=78fb45e862875c072b8cad7555dd54ec6eac2ad0 submitted by /u/Wononscopomuck [link] [comments]”
“They use whatever search API you give them (I'm using a self-hosted SearXNG instance), and camofox prevents websites from being able to tell it's a bot.”
“That;s interesting. I experienced this with other programs to a degree. It's almost like AI can be a mentor now.”
“Lol that screenshot. Completely irrelevant but curious: did you generate that hand into there or did you actually use some image editing program? (dont laugh at me for asking this plz)”
“…And when it does, it takes like 5-10 minutes to generate one image. It’s been acting wacky since last night, where I’m hearing that it may have been down for a while? Is anyone else experiencing this? My page will either go completely blank after it starts to generate an image, and then the page refreshes. This is on desktop and on my iPhone. To be clear, these are not photos that are violating the rules or anything, it’s just been struggling to generate photos since last night. submitted”
“Full credit goes to the artist Masawo Koga ( 古賀マサヲ ). Please check their work out! Prompt: Create an original image of a single character in a 2D illustrated style, with no clear sense of composition. The character should feel like an original late-1980s sci-fi fantasy anime/illustration subject: a dramatic armored warrior with retro futuristic design cues, expressive linework, metallic details, and a slightly pulpy 1988 Japanese illustration sensibility. Use slight motion blur and mild overexpo”
“it keeps posting nothing at all or saying image generation failed submitted by /u/Hexxegone [link] [comments]”
“https://preview.redd.it/gjylybgpnedh1.jpeg?width=1086&format=pjpg&auto=webp&s=e627122ac0ca75ff960348b61c929423bb4e2d12”
“I'm still really impressed by what ChatGPT Image 2 can do with frame animation sheets. You can start from your own uploaded image or just a text idea, and it does a surprisingly good job of understanding how movement should progress across multiple frames. It creates a sequence of individual frames, much closer to traditional animation, where motion is implied by the progression between frames. I made a Custom GPT that focuses on prompting consistent animation sheets, together with a small P”
“https://preview.redd.it/aice8rnldcdh1.jpeg?width=1536&format=pjpg&auto=webp&s=32fb404feadf33210d766e0c65aab058ca5616d3 (Prompt: "Make a photo of the White House but modded to be a proper gaudy and kitsch McMansion with maximum dictator vibes and lots of extensions like a plane hanger as a garage)”
“Out of curiosity, I gave the same prompt to all 4 models (grok-4.5 / gemini-3.5-flash / gpt-5.6-terra / gpt-5.6-sol) to make a product video for Stripe.com. Same URL, same assets — but totally different pacing, motion style, and script choices. Honestly couldn't have predicted which model made which. Video's attached. Curious which one you'd actually use for your own launch — and if anyone else building on these models sees the same variance. submitted by /u/zzJoeyyy [lin”
“Not a huge area that I need to cover and it doesn’t need to be “perfect“, just passable. submitted by /u/Rushinwind [link] [comments]”
“https://preview.redd.it/tkh5cljsp8dh1.jpeg?width=1536&format=pjpg&auto=webp&s=4c4b353f6e3e8b58bfd58a520ca346c821e94ee2”
Niiice !! What are you guys thinking would be the use case ?
“Share link: https://chatgpt.com/share/6a539783-08fc-83ea-bba1-48765fa02e67 Reproduction steps, any user with cross thread memory enabled: upload an image to a new thread, with a prompt similar to "Mark this thread as {threadname}. Confirm i've uploaded an image, reply only with yes or no" in a new thread after this, repeat the 2 prompts included at the top of the shared link. The 2nd response will include the key finding. Cliffs: this demonstrates a situation where a permanent, sil”
“This is the dual-sref finding: the same Witness image used as a plain image prompt caused staging collapse (~12% success, contact drift back), but layered as a second --sref value alongside a numeric style code, it transferred cleanly (16/16 gesture originally, 15/16 staging) with zero compositional contamination. Strong hook because it's counterintuitive — same file, same subject, wildly different result based on mechanism , not content. submitted by /u/jeffbradshaw [link]”
“https://preview.redd.it/c33qh8e545dh1.png?width=1536&format=png&auto=webp&s=ee9f93ce2a6026b4e0c65e0aa704a0c3acfb04e1 Huh. Well, I guess I am ok with that”
