Mastering ChatGPT (OpenAI)Files, data, images and voice · Lesson 7 of 19

Image generation and vision

Article · 15 min · 9 min lecture

Video lecture

Image generation and vision

15 chapters · about 9 min · full transcript

Coming soon

Chapter 1 of 15

Images in ChatGPT

  • Vision analyses
  • Generation creates
  • Use both responsibly

The narrated lecture is in production

Every chapter is scripted and ready. Browse the chapters and read the full transcript now — the video will appear here when it’s published.

Chapters

Two different capabilities

  • Vision means ChatGPT analyses images you give it: screenshots, photos, charts, handwriting, product labels, design drafts.
  • Image generation means ChatGPT creates new images from descriptions (and edits existing ones). ChatGPT's image generation is powered by OpenAI's GPT Image models. In September 2026 OpenAI released ChatGPT Images 2.5, which added a Sketch feature for turning your own drawings into finished images and reduced generation time. The older standalone DALL·E GPT was retired at the end of August 2026; image creation now lives directly in ChatGPT and its images library.

Availability, limits and quality settings vary by plan.

Vision: the describe-first pattern

1. Describe everything you see in this image: text, numbers, labels, colours, layout.
   Don't interpret yet.
2. Using only that description, answer: [your question].
3. List anything you were unsure about reading.

This separates perception from reasoning, so a misread number is caught before it becomes a conclusion. Crop to what matters, use legible resolution, and label multiple images ("Image 1 = before, Image 2 = after").

Image generation: the prompt formula

A strong image prompt covers:

ElementExample
Subjecta ceramic coffee cup with Arabic calligraphy
Settingon a marble café counter in Old Dubai at golden hour
Styleeditorial product photography, natural light
Compositioncup on the left third, clean space at the top for headline text
Lighting and moodwarm, soft shadows, inviting
Format4:5 vertical for Instagram feed
Text (if any)exact words in quotes, and check spelling in the result

Iterate one change at a time ("same image, but cooler morning light") and keep a prompt log so you can reproduce what worked. Use Sketch when you know the composition you want: draw rough shapes and let the model render them.

Editing images

You can upload an image and ask for edits (change the background, remove an object, adjust the style), or select part of a generated image and describe the change. Always compare the edit to the original for unintended changes, especially to faces, logos and product details.

Responsible use

  • Do not misrepresent real products. Use AI images for concepts, mood boards and backgrounds; use real photos of the actual product in ads, or composite real product shots into AI backgrounds.
  • People and likeness: do not create images that impersonate real people or imply endorsements they did not give. Respect OpenAI's usage policies.
  • Disclosure and provenance: OpenAI adds C2PA provenance metadata to ChatGPT-generated images; platforms such as Meta, TikTok and YouTube have their own AI-content labelling rules. Follow them, and your advertising codes.
  • Intellectual property: avoid prompts that copy protected characters, logos or a living artist's distinctive work; check your organisation's policy for commercial use.
  • Check details: hands, text, logos, reflections and product features often contain errors.

Worked example: a restaurant launch

A Karachi restaurant launching a new biryani platter uses ChatGPT to generate three mood-board images for the menu design and social teaser (clearly concept art), and Sketch to test a flat-lay layout. For the launch ads, the team photographs the real dish in the layout ChatGPT helped design, and uses ChatGPT's vision to review the ad drafts for text legibility on mobile. No customer ever sees an AI image presented as the real food.

Hands-on

  1. Use the prompt formula to create one concept image for your brand or project.
  2. Refine it three times, changing one element each time; log what changed.
  3. Upload a screenshot of an ad or landing page (with personal data cropped) and run the describe-first pattern.

A reusable creative-review prompt (vision)

You are reviewing social ad creatives for [brand]. For each attached image:
1. Describe layout, headline text, CTA, colours and any small print.
2. Score 1-5: offer clarity at thumbnail size, brand fit, visual hierarchy,
   mobile legibility.
3. Flag any claim that needs evidence and any missing ad disclosure label.
4. Suggest one concrete fix.
End with a ranking table and "Uncertain readings".

Treat scores as hypotheses to test with real audiences, not results.

Pitfalls

  • Changing five things per iteration and losing what worked.
  • Publishing AI images with garbled text or extra fingers.
  • Uploading screenshots with customer emails or account details visible.

How to measure success

Concept rounds that used to take days take hours, no AI image misrepresents a real product or person, and vision reviews catch issues before publication.

Key takeaways

  • Vision analyses images you upload; image generation (GPT Image, Images 2.5 with Sketch since September 2026) creates and edits images.
  • Describe subject, setting, style, composition, lighting, mood, format and exact text; iterate one change at a time with a prompt log.
  • Don't use AI images to misrepresent real products or people; follow provenance, platform labelling and advertising rules.
  • Check hands, text, logos and product details, and crop personal data from screenshots before uploading.

Check your understanding

Quick questions to lock in the lesson. They don’t count towards your certificate.

  1. Which prompt element helps leave room for a text overlay?
  2. A restaurant wants to advertise a new dish. Which use of AI images is most responsible?
  3. What should you do before uploading a screenshot for vision analysis?

Put it into practice

Use the image prompt formula to create one concept image for your brand or project. Refine it three times and note which change had the biggest effect.

Enrol for free to save your progress

Reading is always free. Enrol to keep your place, take the final assessment and earn a verifiable certificate.